アラビア語翻訳の誤り区間を検出する表層タグ付け
TTLab at AlexandriaX-2026: A Fine-Tuned Surface Tagger for Arabic Machine-Translation Error-Span Detection and Classification
この論文をやさしく読む
ひとことで言うと
アラビア語への機械翻訳で、誤りがある文字列の範囲と種類を見つけるモデル。
何に役立つ?
翻訳の品質評価で誤りの場所を正確に特定する方法の比較材料になる。
この研究の面白いところ
文字位置を保つトークン分類と、ラベルの偏りに対応した損失・方言別しきい値を組み合わせた。
どこまで分かった?
まれな誤りの種類の分類は難しいままと報告されている。数値は競技の開発・試験データでの結果。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
著者らはAlexandriaX-2026のサブタスク3である、アラビア語機械翻訳の誤り区間の検出と分類へのTTLabの提出システムを示す。課題を表層形に対するトークン単位の分類として捉え、評価指標と区間を正確に合わせるため文字位置のオフセットを保持する。ラベルの偏りが大きいため、クラスの重み付けを加えた焦点損失と、方言ごとの復号しきい値を使う。アラビア語の事前学習済みエンコーダー6種類の中では、MARBERTv2が開発データで40.8、試験データで40.91という全体で最良の性能を得て、参加チーム全体で3位となった。誤りの区間を見つける性能は高かった一方、まれな誤りの種類の分類は依然として難しく、件数の少ないカテゴリに対するデータ拡張が必要である。コードは公開されている。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-24(UTC)
- 最新改訂
- 2026-09-24 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-24 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
We present TTLab's submission to the AlexandriaX-2026 Subtask~3 on Arabic MT error span detection and classification. Our system frames the task as token-level classification over surface forms, preserving character offsets to ensure exact alignment with the evaluation metric. To handle severe label imbalance, we employ a focal loss with class weighting and dialect-specific decoding thresholds. Among six Arabic pre-trained encoders, MARBERTv2 achieves the best overall performance of 40.8 and 40.91 on the development and test set, respectively, ranking $\nth{3}$ out of all participating teams. While our system localizes error spans effectively, classification of rare error types remains challenging, highlighting the need for data augmentation for tail categories. The code is available at ${\href{https://github.com/ENTAILab/arabic-dialectal-mt-error-span-detection}{\faGithub~ TTLab at AlexandriaX-2026}$
著者のコメント
Accepted at ArabicNLP 2026, shared task AlexandriaX-2026
arXiv ID: 2609.29633 / 要約の誤りについて