arXiv論文メモ
新着一覧
cs.LG · 査読状況未確認

文章と構造の矛盾を使ってSNSのボットを検出する

CSC: Calibrated Simplicity for Conflict-Aware Social Bot Detection in the LLM Era

Yipeng Qian, Pengjie Zhao, Chaoxi Niu

この論文をやさしく読む

ひとことで言うと

SNSの投稿文が人間らしくても、つながりやプロフィールとの矛盾からボットを見つける方法を提案した。

何に役立つ?

考えられる用途は、SNS上のボット検出で文章・グラフなどの判断を較正して組み合わせること。

この研究の面白いところ

文章を人間の文に入れ替える試験で、文章単独とグラフを併用した判断の違いを示した点。

どこまで分かった?

実験は挙げられた三データセットと偽装試験の条件に基づく。実運用での誤判定率や攻撃者の適応後の性能は要旨にない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

SNSのボット検出は、誤情報の拡散、連携した操作、公共の議論のゆがみからプラットフォームを守るために重要である。しかし大規模言語モデルにより、自然で大量の文章を安く作れるため、文章だけではボットを見分けにくくなった。アカウントの文章は人間らしくても、つながりの構造、プロフィールの属性、異なる情報間の整合性には疑わしさが残ることがある。最近のグラフ型検出器は、少数の典型例の選択、適応的なゲート、構造に固有の制御など複雑な仕組みを加えるが、本研究の実験は、複雑さだけでは矛盾への対処が安定しないことを示す。 そこで、矛盾を考慮し、較正を重視する簡潔な枠組みCSCを提案する。構造上の有用な偏りを残しつつ不安定な規則を取り除いた典型例に基づく簡潔なグラフ専門器、異なるモデルの確信度を合わせてから後段で統合する単体制約付きの較正、情報源間の不一致を扱う軽量な専門器を組み合わせる。TwiBot-22、TwiBot-20、MGStBot-largeでの実験では、ほかのベンチマークでも競争力を保ちながら、較正した運用点での判断品質が改善した。追加の解析では、較正は確信度の信頼性を高め、不一致専門器は主に矛盾が大きいか判定境界に近い領域を局所的に修正し、グラフ側の制御を簡潔にすると安定性と費用の兼ね合いが良くなった。文章の偽装を狙った試験では、一部のボット文章を対応する人間の文章に置き換えると、文章だけの判定器の性能は大きく低下したが、均衡を取った試験集合ではグラフと統合した証拠は安定していた。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-20(UTC)
最新改訂
2026-09-20 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Social bot detection is essential for protecting online platforms from misinformation amplification, coordinated manipulation, and distorted public discourse. However, large language models have made social bots much harder to detect from text alone because semantic camouflage is now cheap, fluent, and scalable. The resulting challenge is modality conflict: an account may look human-like in semantics while remaining suspicious in graph structure, profile attributes, or cross-modal consistency. Recent graph-based detectors tackle this limitation by adding graph-side complexity, such as sparse prototype selection, adaptive gating, or architecture-specific control logic, yet our experiments suggest that complexity alone is not the most reliable way to resolve such conflict. We therefore propose CSC, a calibrated-simplicity framework for conflict-aware LLM-era social bot detection. The framework combines three design choices: a simplified prototype-guided graph expert that retains useful structural biases while removing unstable graph-side heuristics, calibrated simplex-constrained fusion that aligns heterogeneous confidence spaces before late fusion, and a lightweight inconsistency expert that models cross-modal disagreement. Experiments on TwiBot-22, TwiBot-20, and MGStBot-large show that \textsc{CSC} improves calibrated operating-point decision quality while remaining competitive across external benchmarks. Further analyses show that calibration improves confidence reliability, the inconsistency expert mainly provides localized corrections in high-conflict or near-threshold regions, and simplified graph-side control yields a better stability-cost trade-off. A targeted semantic-camouflage stress test further shows that replacing selected bot text with matched human text sharply degrades the standalone text expert while leaving graph and fused evidence stable on a balanced challenge set.

著者のコメント

Accepted by CIKM 2027

arXiv ID: 2609.23320 / 要約の誤りについて