arXiv論文メモ
新着一覧
cs.IR · 査読状況未確認

問い合わせ側だけを更新する文書画像検索の適応

Test-Time Adaptation with Query-Dependent Residuals for Visual Document Retrieval

Zeliang Li, Xiaofen Xing, Kailing Guo, Xiangmin Xu

この論文をやさしく読む

ひとことで言うと

既存の文書埋め込みを変えず、問い合わせの表現を再順位付け情報から改善する。

何に役立つ?

考えられる用途は、文書索引を作り直せない状況での画像ページ検索の改善である。

この研究の面白いところ

再順位付けされなかったページも、全索引上の学生分布を通して学習に関わる。

どこまで分かった?

八課題と五つの基盤モデルでの比較を報告する。要旨には個別の改善幅はない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

文書画像検索は、運用前に計算したページの埋め込みに依存するため、エンコーダーのパラメータを変えられず文書全体を再符号化できない場合、適応が難しい。再順位付け器は関連性の信号を与えるが、通常は選ばれた問い合わせと候補ページにしか適用されない。Q-REACTは、限られた再順位付けのフィードバックを再利用できる検索改善に変える、問い合わせ側のテスト時適応法である。共通の低ランク変換を学び、問い合わせに依存した残差を作り、適応した問い合わせスコアと文書レベルの文脈を組み合わせる。さらに、再順位付け器の選好を、課題固有の全ページ索引で正規化した学生分布に蒸留する。これにより、採点されなかったページも保存済み埋め込みで競争に参加でき、エンコーダーとページ索引は固定したままにできる。ViDoRe V3の八課題と、公開・非公開の五つの基盤モデルで、Q-REACTは疎な評価予算と全件を覆う予算の双方で比較手法より平均検索性能が高かった。また、保留した問い合わせや課題にも転移し、推論時の追加負荷は小さかった。限られた再順位付けの情報を、検索器の再学習や再構築なしで問い合わせ群に広く活用できることを示す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-23(UTC)
最新改訂
2026-09-23 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Visual document retrieval (VDR) systems depend on page embeddings computed before deployment, which makes adaptation difficult when encoder parameters or corpus re-encoding are unavailable. Rerankers provide useful relevance signals, but conventional reranking applies them only to selected queries and candidate pages. We introduce Q-REACT, a query-side test-time adaptation method that converts limited reranker feedback into reusable retrieval improvements. Q-REACT learns a shared low-rank transformation that produces query-dependent residuals, combines adapted query scores with document-level context, and distills reranker preferences with a student distribution normalized over the complete task-specific page index. This design lets unscored pages compete through cached embeddings while keeping the encoders and page index fixed. Across eight ViDoRe V3 tasks and five open-weight and proprietary backbones, Q-REACT improves average retrieval over evaluated baselines at sparse and full-coverage budgets, transfers to held-out queries and tasks, and adds little inference overhead. The results show that finite reranker feedback can be amortized across a query collection without retraining or rebuilding the retriever.

arXiv ID: 2609.27688 / 要約の誤りについて