arXiv論文メモ
新着一覧
cs.AI · 査読状況未確認

長いエージェント履歴を根拠を残して圧縮

Stable Geometry with Divergent Task Evidence for Efficient Long-Horizon Agent Compression

Mingxuan Wang, Fei Luo, Bo Wang, Guorun Yao, Yinglong Guo, Chao Ning, Hongyue Chen, Yanbiao Ma, Jungong Han

この論文をやさしく読む

ひとことで言うと

エージェントの長い作業履歴から、次の行動に必要な証拠を優先して残す圧縮方法です。

何に役立つ?

長期間動くエージェントの文脈長と推論費用を減らす用途が考えられる。

この研究の面白いところ

見た目の幾何学的な類似度が0.98でも、次の行動に必要な情報の保持率は大きく変わることを示した。

どこまで分かった?

トークン使用量と報酬の結果は要旨で扱った課題での評価であり、すべての種類のエージェント履歴に対する保証ではない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

長期間動くエージェントには履歴が積み重なり、文脈の長さと推論費用が増える。著者らは、幾何学的な重複だけでは安全な圧縮の基準にならないことを見いだした。エージェントの履歴には強い低次元構造があるものの、全体の幾何学的な形が似ていても、課題の根拠がどれだけ残るかは大きく異なる。同じ数のブロックを保持する条件では、根拠を考慮した選択により、次の行動の上位3候補が残る率が0.31から0.69へ上がった一方、重心の類似度は0.98のままだった。制御した置き換え実験でも、全体の幾何学的な指標がほぼ変わらないまま、行動に関係する情報が大きく変わり得ることを示した。この幾何と根拠のずれを踏まえ、学習不要の圧縮法Geometry Guided Evidence Preserving Memory(GEM)を導入する。課題と実行に関わる根拠を先に保護し、その後で幾何学的な残差を使って残りを覆う。GEMは課題当たりの平均合計トークン使用量を269万から211万へ、21.4%減らし、課題報酬は同程度に保った。結果は、効率的な履歴圧縮では幾何学的な網羅だけでなく、課題の根拠を保つことを最適化すべきだと示す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-23(UTC)
最新改訂
2026-09-23 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Long horizon agents accumulate growing interaction histories that increase context and inference costs. We find that geometric redundancy alone is an insufficient criterion for safe compression. Although agent histories exhibit strong low dimensional structure, similar global geometry can preserve very different amounts of task evidence. At identical retained block counts, evidence aware selection raises next action Top 3 retention from 0.31 to 0.69, while centroid similarity remains 0.98. Controlled replacement further shows that action related information can be substantially altered while global geometric measures remain nearly unchanged. Motivated by this gap between geometry and evidence, we introduce Geometry Guided Evidence Preserving Memory (GEM), a training free compressor that protects task and execution evidence before using geometric residuals to complete coverage. GEM reduces mean combined token usage from 2.69M to 2.11M per task, a 21.4% reduction, while maintaining comparable task reward. Our results show that efficient agent history compression should optimize for preserved task evidence rather than geometric coverage alone.

arXiv ID: 2609.27332 / 要約の誤りについて