17個の言語特徴で文章間の含意を説明可能に判定
Linguistic Features for Interpretable Textual Entailment
この論文をやさしく読む
ひとことで言うと
文の意味関係や否定、埋め込みの情報変化を17個の特徴にまとめ、文章同士の含意・中立・矛盾を説明しやすいモデルで判定します。
何に役立つ?
文章の推論判定で、どの言語的手掛かりが結果を支えたか調べたい場合に役立ちます。比較的小さいモデルでも評価データ上で高い正解率を得ています。
この研究の面白いところ
意味関係を明示的に扱う特徴が中心でありながら、分布的な情報も中立や矛盾の検出を補うという役割分担を分析しています。
どこまで分かった?
正解率はSICKとSICK-CEでの結果です。計算量が小さいとの記述はありますが、要旨には具体的な実行時間や消費電力はなく、他言語・他領域への一般化も示されていません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
自然言語処理におけるニューラルモデルは成功を収めているものの、ブラックボックスとしての性質が解釈可能性を制限し、予測を支える言語現象を見えにくくしている。本研究では、テキスト含意認識のための説明可能なハイブリッドモデルSLITEを提示する。これは、合成的な実体間の意味的な適合・不適合に基づく構造・関係層と、前提文と仮説文の埋め込み表現間での情報変化の構造的パターンに基づく分布・情報層という、相補的な二つの意味解析層を統合する。 実体単位の意味関係、肯定・否定の極性を考慮した語彙照合、類似度行列の意味的な部分表現の対応付け尺度を組み合わせた17個の特徴を提案する。これにはエントロピーや移動エントロピーに基づく尺度も含まれる。これらの特徴で学習したロジスティック回帰は、3クラスのSICKで正解率83%、SICK-CEで96%を達成する。IsoLexを4パーセントポイント上回り、RoBERTaの計算複雑性の一部で、その正解率との差は2パーセントポイント以内に収まる。 アブレーション実験とSHAP解析から、構造・関係特徴が分類の主要な要因である一方、分布・情報特徴も、とくに中立や矛盾の検出で不可欠な補完的寄与をすることを確認する。結果は、ハイブリッド手法のさらなる探究が、巨大なニューラル構造に対する実行可能で科学的に有意義な代替となることを示している。言語理論と推論の計算モデルとの対話を強化することを期待する。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-21(UTC)
- 最新改訂
- 2026-09-21 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-21 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Despite the success of neural models in natural language processing, their black-box nature limits interpretability and conceals the linguistic phenomena underlying their predictions. We present SLITE, an explainable hybrid model for Recognizing Textual Entailment that integrates two complementary layers of semantic analysis: a structural-relational layer, based on semantic compatibility and incompatibility between compositional entities, and a distributional-informational layer, based on structured patterns of information change between embedding-based representations of the premise and the hypothesis. We propose 17 features that combine entity-level semantic relations, polarity-sensitive lexical matching, and alignment measures over semantic sub-representations of the similarity matrix, including measures based on entropy and transfer entropy. A logistic regression trained on these features achieves an accuracy of 83% on three-class SICK and 96% on SICK-CE, outperforming IsoLex by 4 percentage points and falling within 2 percentage points of RoBERTa with a fraction of its computational complexity. Ablation studies and SHAP analysis confirm that structural-relational features are the primary drivers of classification, while distributional-informational features provide essential complementary contributions, particularly for detecting neutrality and contradiction. Our results demonstrate that further exploration of hybrid approaches is a viable and scientifically productive alternative to massive neural architectures, and we hope they will strengthen the dialogue between linguistic theory and computational modeling of inference
著者のコメント
38 pages, 5 figures, 8 tables
arXiv ID: 2609.24932 / 要約の誤りについて