健康動画の入り組んだ主張を細かく構造化する
Structured Claim-Level Discourse Representations for Dense Health Narratives
この論文をやさしく読む
ひとことで言うと
健康を扱う動画を話題単位でまとめるだけでなく、個々の主張の立場や語り方まで分けて分析します。
何に役立つ?
健康情報がどのように提示され説得力を持たせられているかを調べるための評価基盤になります。
この研究の面白いところ
1分あたり平均13.22の主張という密度を捉え、内容の分類と、根拠や修辞などの語用論的な属性を区別しています。
どこまで分かった?
データは4領域60動画の1,191主張です。主張の医学的な真偽を判定する実証ではなく、高次元の語用論的分析は現行モデルが苦手と報告しています。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
ソーシャルメディアの動画における健康の言説は、短い会話の中に、複数のテーマ側面、立場、根拠の示し方、修辞的な機能にまたがる主張が密接に絡み合っていることが多い。既存の方法は主に粗い話題単位、感情、立場に基づく表現に依存し、この構造を十分に捉えていない。本研究の分析では1分あたり平均13.22個の原子的な主張を特定し、より豊かな主張単位の言説表現の必要性を示す。密度の高い健康の語りにおける、主張単位の言説分析の構造化枠組みを導入する。原子的な主張を、テーマ側面、立場、多次元の語用論的な言説属性と結びつける組によって、言説をモデル化する。 この設定を支えるため、4つの健康領域にまたがり、60本の動画から手作業で注釈した1,191の主張を含むベンチマークを構築する。この枠組みを用い、異なる言説文脈の設定で、自動的な構造化言説分析を評価する。結果は、現行LLMがテーマ分類と立場予測で高い性能を示す一方、高次元の語用論的な特徴づけには苦戦することを示す。また、言説課題によって有効な文脈推論の形が異なることも分かり、将来のシステムには課題の分解と特化した推論戦略が必要となる可能性が示唆される。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-16(UTC)
- 最新改訂
- 2026-09-16 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-16 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Health discourse in social media videos often contains densely entangled claims spanning multiple thematic aspects, stances, evidential frames, and rhetorical functions within short conversational spans. Existing approaches largely rely on coarse topic-level, sentiment-based, or stance-oriented representations that do not adequately capture this structure. Our analysis identifies an average of 13.22 atomic claims per minute, motivating richer claim-level discourse representations. We introduce a structured framework for claim-level discourse analysis in dense health narratives. Our framework models discourse through tuples linking atomic claims with thematic aspects, stance, and multidimensional pragmatic discourse attributes. To support this setting, we construct a benchmark spanning four health domains with 1,191 manually annotated claims from 60 videos. Using this framework, we evaluate automated structured discourse analysis under different discourse context settings. Results show that current LLMs achieve strong performance on thematic categorization and stance prediction, but struggle with high-dimensional pragmatic profiling. We also find that different discourse tasks benefit from different forms of contextual reasoning, suggesting that future systems may require task decomposition and specialized inference strategies.
arXiv ID: 2609.18905 / 要約の誤りについて