タグを保ち自然に訳すための構造付き文章翻訳
Tag-Aware Structured Text Translation: Towards a Systematic Understanding
この論文をやさしく読む
ひとことで言うと
書式タグを壊さずに文章を自然に翻訳するため、データ作成から学習までまとめて改善する研究です。
何に役立つ?
ウェブ文章やタグ付き文書を翻訳するとき、構造の保持と訳文の自然さを両立する設計に役立ちます。
この研究の面白いところ
タグの多様性と自然さの両立をデータ合成で扱い、四つの小課題と三つの報酬でも個別に能力を伸ばしています。
どこまで分かった?
評価は要旨で挙げた六つの言語方向での結果です。ほかの言語やタグ形式で同じ改善が得られるとは示されていません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
インターネット上の文章には、構造、意味、機能を担う書式タグが多く含まれる。現在の大規模言語モデルによる翻訳システムは、タグ付き文章を扱う際に、自然な訳文とタグの忠実な保持を両立させることに苦労している。この両立には、データ合成、能力の獲得、複数の目的への適合という、相互につながる三段階で体系的に取り組む必要があると主張する。データ段階では、合成データの生成においてタグ構造の多様性と訳文の自然さとの間に根本的なトレードオフがあることを特定し、形式化する。既存の方法は一方を改善して他方を損なう。そこで、二つの大規模言語モデルによるタグ付きデータ合成法を組み合わせ、多様で自然なデータを作る混合合成法Hy-LSTを提案する。能力の段階では、タグを考慮した翻訳を難しさの異なる四つの小課題に分け、複数課題による教師あり追加学習の枠組みで、狙った能力の獲得と知識の移行を可能にする。適合の段階では、グループ相対方策最適化の枠組みの中で、流暢さ、タグの忠実さ、タグで区切られた範囲内の訳の品質をそれぞれ狙う三つの報酬関数を設計する。これらを共同で最適化すると、一つの報酬だけを使う方法を一貫して上回った。英語から中国語、日本語、ドイツ語、フランス語、ロシア語、およびドイツ語からフランス語への六方向の実験では、三段階それぞれが測定可能な改善に寄与し、全体の仕組みは既存手法を有意に上回った。定性的な分析でも、具体的な誤りの型と、提案手法による学習後のその緩和が示された。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-24(UTC)
- 最新改訂
- 2026-09-24 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-24 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Internet texts are replete with format tags that carry structural, semantic, and functional meaning. Current large language model (LLM)-based translation systems struggle to balance translation fluency with tag fidelity when processing tagged text. We argue that resolving this tension requires a systematic approach at three interconnected levels: data synthesis, capability building, and multi-objective alignment. At the data level, we identify and formalize a fundamental trade-off between structural tag diversity and translation naturalness in synthetic data generation; existing methods optimize for one at the expense of the other. We propose a hybrid synthesis strategy (Hy-LST) combining LLM-based synthesis tag method and Two-Stage LLM-based synthesis tag method to produce both diverse and natural tagged data. At the capability level, we decompose tag-aware translation into four sub-tasks of increasing difficulty in a multi-task supervised fine-tuning framework, enabling targeted capability acquisition and knowledge transfer. At the alignment level, we design three complementary reward functions under a group relative policy optimization framework, each targeting a distinct objective (fluency, tag fidelity, and tag-scoped translation quality), and show that joint optimization consistently outperforms single-reward alternatives. Experiments on six language directions (en2zh, en2ja, en2de, en2fr, en2ru, de2fr) demonstrate that each level contributes measurable improvements, and the complete system significantly outperforms existing methods. Qualitative analysis reveals specific error patterns and their mitigation after training with our method.
arXiv ID: 2609.29131 / 要約の誤りについて