arXiv論文メモ
新着一覧
physics.soc-ph · 査読状況未確認

ベンガル語の商品推薦対話データBanglaShop-CRSを作成

BanglaShop-CRS: A User-Centric Bangla Dataset for Conversational Recommendation

Tabia Tanzin Prama, Christopher M. Danforth, and Peter Sheridan Dodds

この論文をやさしく読む

ひとことで言うと

実際の購入履歴やレビューを基に、ベンガル語で商品を相談する対話を人工的に作ったデータセットです。

何に役立つ?

英語中心だった対話型推薦の学習と評価を、ベンガル語や英語混在の利用場面へ広げるための基盤になります。

この研究の面白いところ

会話の自然さだけでなく、正しい利用者記録とシャッフルした記録を比較して、対話が元の好みに結び付いているかを評価しています。

どこまで分かった?

実際の行動データに基づきますが、対話自体は合成データです。人手評価は母語話者5人によるもので、実サービスでの購買満足度や長期的な推薦効果を測った結果ではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

対話型推薦システム(CRS)では、利用者が自然言語でやり取りしながら好み、制約、フィードバックを伝えられる。しかし既存のCRS資源は、英語などの資源が豊富な言語に集中し、ベンガル語やベンガル語と英語が混在する場面は十分に扱われていない。この不足を補うため、実際の電子商取引での行動を基にした、大規模で利用者中心の合成ベンガル語商品推薦対話データセットBanglaShop-CRSを導入する。10の商品領域にわたり、27,178件の複数ターン対話、274,802発話、360万トークンを含む。生成処理には、利用者の購入履歴、肯定的・否定的フィードバック、レビュー文を組み込み、対話内容と利用者の好みの整合性を保つ。 カタログ制約付きと自由語彙の推薦プロトコルのもとでBanglaShop-CRSを評価する。結果は、対話の文脈が推薦品質を改善し、ファインチューニングにより各モデルでさらに改善することを示す。5人のベンガル語母語話者による人手評価では、対話の流暢さ、情報量、論理性、一貫性が確認され、すべての評価次元で有意な一致が得られた。事実への根拠付けの評価でも、シャッフルした利用者記録より正しい記録との強い整合性が示され、評価者間には相当程度の一致(κ=0.65)があり、人とGPT-5.1の判断も同程度だった。BanglaShop-CRSは、ベンガル語の対話型推薦を進めるための、規模を拡張可能なベンチマークを提供する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-16(UTC)
最新改訂
2026-09-16 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Conversational recommender systems~(CRS) enable users to express preferences, constraints, and feedback through natural language interaction. However, existing CRS resources are concentrated in English and other high-resource languages, leaving Bangla and code-mixed Bangla--English settings underrepresented. To address this gap, we introduce BanglaShop-CRS, a large-scale user-centric synthetic Bangla conversational recommendation dataset grounded in real e-commerce behavior. It contains 27,178 multi-turn dialogues, 274,802 utterances, and 3.6M tokens across 10 product domains. Our generation pipeline incorporates user purchase histories, positive and negative feedback, and review texts to maintain consistency between dialogue content and user preferences. We evaluate BanglaShop-CRS under catalog-constrained and open-vocabulary recommendation protocols. Results show that dialogue context improves recommendation quality, while fine-tuning yields further gains across models. Human evaluation by five native Bangla-speaking annotators confirms the fluency, informativeness, logicality, and coherence of the dialogues, with significant agreement across all dimensions. Factual-grounding evaluation further shows stronger alignment with correct than shuffled user records, with substantial inter-annotator agreement ($\kappa=0.65$) and comparable human and GPT-5.1 judgments. BanglaShop-CRS provides a scalable benchmark for advancing conversational recommendation in Bangla.

arXiv ID: 2609.18715 / 要約の誤りについて