arXiv論文メモ
新着一覧
cs.IR · 査読状況未確認

配信推薦で大規模言語モデルの利用者プロファイルはいつ有効か

When LLM-Based User Profiling Adds Value in Production Streaming Recommendation

Milad Sabouri, Neeraj Sharma, Sardar Hamidian, Shaghayegh Agah

この論文をやさしく読む

ひとことで言うと

推薦システムの利用者表現について、LLMによる好みの文章化と埋め込みの集約を、時間の扱いも含めて比較した。

何に役立つ?

LLMによる高コストなプロファイル生成を導入するか判断する際、利用者の行動類型や評価指標、時間窓を分けて比べる視点を提供する。

この研究の面白いところ

表現方法と直近・過去の行動の分離を掛け合わせ、四つの方法を本番データで比較している。

どこまで分かった?

要旨は比較で違いが現れたと述べるが、具体的な精度差や費用対効果の数値は示していない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

個人向け推薦の質は、過去の行動から利用者の表現をどう構築するかに大きく依存する。内容に基づく推薦の意味的な利用者プロファイルには二つの方法がある。一つは、項目の意味埋め込みを数値的に集約する方法である。もう一つは、大規模言語モデル(LLM)で好みを自然言語に要約し、テキストエンコーダで表現する方法である。どちらも直近の行動とそれ以前の行動を時間的に分けて扱える。LLMによるプロファイル生成は集約法より大幅に高価なため、その追加費用がいつ正当化されるかが問題になる。本研究は、表現の種類と時間の扱いを要因として組み合わせた四つの意味的なプロファイル作成法を、本番環境の実データで体系的に比較した。その比較から、利用者の行動類型ごとの違い、推薦の精度と精度以外の質の両面での違い、直近と過去を分ける時間窓の設定による違いを明らかにした。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-23(UTC)
最新改訂
2026-09-23 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Personalized recommendation depends critically on how user representations are constructed from historical behavior. Two paradigms have emerged for constructing semantic user profiles in content-based recommendation. First, aggregate methods derive user representations as numerical aggregates of semantic item embeddings. Second, LLM-based methods generate natural-language summaries of user preferences and encode them through a text encoder. Each paradigm can be combined with temporal disentanglement of recent versus historical behavior. LLM-based profile generation is significantly more expensive than aggregate approaches, raising the question of when this additional cost is justified. We present a systematic comparison of four semantic user-profiling strategies, factorially crossed across representation type and temporal handling, evaluated on a real-world production dataset. The comparison reveals how these strategies differ across user behavior types, across both accuracy and beyond-accuracy dimensions of recommendation quality, and across the temporal-window setting that governs the disentanglement.

arXiv ID: 2609.27183 / 要約の誤りについて