arXiv論文メモ
新着一覧
cs.LG · 査読状況未確認

補助課題の同時学習が汎化を改善する仕組みの理論

An Analytical Theory of Auxiliary Learning

Federico Milanesio, Alessandro Ingrosso, Matteo Osella

この論文をやさしく読む

ひとことで言うと

別の課題も同時に学ぶと目標課題の性能が改善する理由を、数式と数値実験で調べた研究です。

何に役立つ?

補助課題の選び方を、課題間の相関やラベル雑音に基づいて考える理論的な手掛かりになります。

この研究の面白いところ

線形の場合は汎化誤差の式を導き、非線形の場合は主課題・補助課題・単独学習の誤差を結ぶ関係を示しています。

どこまで分かった?

理論は教師・生徒の枠組みと入力次元が大きい極限に基づきます。実際の大規模モデル一般への適用は要旨では検証されていません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

補助学習は、目標課題に加えて別の課題も同時に学習させ、ニューラルネットワークの目標課題での性能を高める最適化の方法である。しかし改善が生じる仕組みは十分に理解されていない。著者らは教師・生徒の枠組みでこの問題を調べ、入力の次元が大きい極限におけるオンライン確率的勾配降下法の動きを記述する、閉じた微分方程式系を導く。 線形ネットワークでは、学習率の主要な次数までの汎化誤差を閉形式で表し、課題間の相関とラベルの雑音が補助学習の利益をどう決めるかを数量化する。非線形の活性化関数については、主課題と補助課題の誤差をそれぞれの単一課題学習時の誤差と結び付ける一般的な関係を示す、揺動散逸に基づく解析理論を作る。数値実験は理論予測を支持し、最適解へ向かう駆動と勾配の雑音との釣り合いを取ることで、補助課題が汎化を改善する仕組みを示している。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Auxiliary learning is an optimization paradigm in which a neural network's performance on a target task is improved by jointly training it on additional tasks. However, the mechanisms behind this improvement remain poorly understood. We study this problem using a teacher-student framework and derive a closed system of differential equations describing the dynamics of online stochastic gradient descent in the large-input limit. For linear networks, we obtain a closed-form expression for the generalization error to leading order in the learning rate, quantifying how task correlations and label noise determine the benefit of auxiliary learning. For non-linear activation functions, we develop a fluctuation-dissipation analytical theory that establishes a general relation linking the main and auxiliary errors to the corresponding single-task error. Numerical experiments support the theoretical predictions and show how auxiliary tasks improve generalization by balancing the forcing dynamics towards the optimal solution with gradient noise.

著者のコメント

Under review as a conference paper

arXiv ID: 2609.29774 / 要約の誤りについて