arXiv論文メモ
新着一覧
cs.CL / cs.LG · 査読状況未確認

投稿者の意図と社会的文脈から陰謀論投稿を見分ける

Agentic Detection of Online Conspiracies

Lior Biton and Oren Tsur

この論文をやさしく読む

ひとことで言うと

陰謀論に関する投稿でも、支持しているのか批判や風刺なのかを、投稿の社会的文脈を調べて判定する手法です。

何に役立つ?

投稿の文面だけでは意図を取り違えやすい場合に、周囲の投稿や社会的文脈を使う分類方法を検討する材料になります。自動判定を運用したときの影響や公平性までは要旨に示されていません。

この研究の面白いところ

約4年間の公開ヘブライ語ツイートの80~90%を含むデータから文脈を復元し、同じ文脈を一括で与えた非エージェント型モデルよりも、必要な証拠を事例ごとに探すエージェント型の方が高い性能を示しました。

どこまで分かった?

評価はヘブライ語ツイートと人手で注釈した難しい事例のデータセットに基づきます。具体的な性能値や他言語・他媒体への一般化は要旨には記載されていません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

ソーシャルメディア上の陰謀論的な言説は、明示的な主張や安定した語句として現れるとは限らない。同じ文面でも、支持、正当な懸念、批判、風刺、嘲笑を表し得る。したがって中心的な課題は、陰謀論に関係する主張を見つけるだけでなく、発話者の意図、すなわち発話行為としての力を推定することにある。本研究は、関連する社会的文脈を用いればこれが可能だと論じ、社会的情報を問い合わせるツール群を備えたエージェント型の枠組みを提案する。 2018年末から2023年初めまでの約4年間に公開されたヘブライ語の公開ツイートの80~90%を含む独自のデータセットで有効性を示す。この期間には複数の選挙や、新型コロナウイルス感染症の流行と関連するワクチン接種運動が含まれる。この広い収集範囲を利用して、さまざまな社会的文脈を復元できる。人手で注釈した、判定が難しい事例を含むデータセットで評価したところ、文脈を考慮する作業手順はテキストだけの分類を一貫して上回った。エージェント型の枠組みは、エージェントと同じ文脈を与えられた非エージェント型モデルを含む、ほかの枠組みや設定よりも有意に高い性能を示した。結果、誤り、効率(トークン使用量)のトレードオフも分析する。これらの知見は、陰謀論検出を社会的文脈の中での解釈課題と捉える考えを支持する。有効な分類には文脈にアクセスできるだけでなく、各事例でその時点の推論に必要な証拠だけをツールで尋ねる、適応的な推論も関わる。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Conspiratorial discourse on social media is not always expressed through explicit claims or stable lexical markers. The same surface content may express endorsement, legitimate concerns, criticism, satire, or mockery. The main challenge is therefore not only recognizing conspiracy-related claims, but inferring the speaker's intent -- the utterance's illocutionary force. We argue that this can be achieved through the use of relevant social contexts and propose an agentic framework, equipped with a set of tools supporting social queries. We demonstrate the benefits of our approach on a unique dataset of Hebrew tweets, covering 80\%--90\% of the public Hebrew tweets published over a four-year span (late 2018-- early 2023), encompassing several election cycles as well as the COVID pandemic years and related vaccination campaigns. This extensive coverage can be used in recovering different social contexts. Evaluating our framework on a manually-annotated adversarial dataset, we find that context-aware workflows consistently outperform text-only classification and that the agentic framework performs significantly better than other frameworks and settings, including a non-agentic model exposed to the same contexts available to the agent. We further provide an analysis of the results, the errors and efficiency (token economy) tradeoffs. These findings support viewing the task of conspiracy detection as a socially embedded interpretation task, in which effective classification depends not only on access to contexts, but also on adaptive reasoning in which the agent uses tools on a per-case basis, asking only for evidence relevant to its current reasoning step.

arXiv ID: 2609.30250 / 要約の誤りについて