arXiv論文メモ
新着一覧
cs.LG / cs.IT / math.IT · 査読状況未確認

関数の反復合成が生む近似能力の限界と高速収束

Neural Approximation by Function Composition: Rigidity and Doubly Exponential Convergence

Wentao Huang and Haizhang Zhang

この論文をやさしく読む

ひとことで言うと

同じ関数を何度も合成する方法で、どこまで関数を近似でき、深さとともに誤差がどれだけ減るかを証明した。

何に役立つ?

深いニューラルネットの近似能力を数学的に理解し、深さの配分を設計するための理論になる。

この研究の面白いところ

区分線形の生成関数には表せる滑らかな関数の限界がある一方、滑らかな生成関数なら平方関数で二重指数的な誤差減少を構成した。

どこまで分かった?

二重指数的な誤差減少は平方関数と任意の固定多項式についての構成である。べき級数には別の誤差率と係数条件、内部領域という条件がある。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

深いニューラルネットワークはアフィン写像と非線形活性化関数の合成で関数を近似するが、合成そのものがどのように近似能力を生むかは十分に理解されていない。著者らは、一つのスカラー生成関数を繰り返し適用したものの、幾何学的な重みを付けた和という基本的な仕組みを調べる。これは、関数x−x²をテント写像で構成する古典的手法や、Yarotsky、W. Eらが深いネットワークの近似能力を解析するのに使った再帰的表現の基礎となる。まず剛性定理を示す。有限個の区間を持つ連続な区分線形生成関数では、この方法で表せるC³級の関数は高々二次関数である。アフィンではない二次関数の場合、幾何学的な係数は少なくとも1/4となる。この結果はテント写像の方法の限界を示し、階層的基底や再帰的多項式の構成を用いた既存手法を補う。次に、正確な余りの恒等式を手掛かりに、反復によって平方関数の近似誤差が全体の深さに対して二重指数的に減る滑らかな生成関数を構築し、乗算モジュールを通して任意の固定した多項式にも拡張する。区間[−1,1]^d上で係数が絶対総和可能なべき級数については、単項式の次数に応じて深さを配分すると、内部の各立方体で一様近似誤差O(exp(−cL^(1/d)))が得られる。これらは生成関数の動きと余りの評価が、深いネットワークの深さ配分と近似速度を支配することを示す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-22(UTC)
最新改訂
2026-09-22 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Deep neural networks approximate functions by composing affine maps with nonlinear activations, but how composition itself creates approximation power is not yet fully understood. We investigate a fundamental mechanism: geometrically weighted sums of iterates of a single scalar generator function. This mechanism underpins the classical tent-map construction of the function \(x - x^2\) and related recursive representations used by Yarotsky, W. E, et al., to analyze the approximation powers of deep neural networks. First, we establish a rigidity theorem: for continuous piecewise linear generators with a finite number of segments, any \(C^3\) function that can be represented in this way is at most quadratic. For non-affine quadratic functions, the geometric factor is at least $1/4$. This result both reveals limitations of the tent-map approach and complements existing methods based on hierarchical bases and recursive polynomial constructions. Second, using an exact remainder identity as guidance, we construct a smooth generator whose iterates yield doubly exponential error decay in total depth for square approximation and, through multiplication modules, for each fixed polynomial. For power series with absolutely summable coefficients on \([-1,1]^d\), distributing depth according to monomial degree yields a uniform approximation error of order \(O(e^{-cL^{1/d}})\) on each interior cube. These findings demonstrate how generator dynamics and remainder estimates govern depth allocation and approximation rates of deep neural networks.

arXiv ID: 2609.25874 / 要約の誤りについて