arXiv論文メモ
新着一覧
eess.AS / cs.SD · 査読状況未確認

多言語の会話音声を扱う第2回MLC-SLM評価課題

The Second MLC-SLM Challenge: Multilingual Conversational Speech Diarization, Recognition, and Understanding

Bingshen Mu, Mingchen Shao, Zhennan Lin, Liumeng Xue, Hexin Liu, Lei Xie, Eng Siong Chng, Longshuai Xiao, Qiangze Feng, Daliang Wang

この論文をやさしく読む

ひとことで言うと

多言語の会話音声を認識・理解するモデルを比較する第2回の評価課題と、参加結果をまとめた報告。

何に役立つ?

公開データ、評価手順、基準システムや参加手法を、今後の会話音声研究の比較基盤として利用できる。

この研究の面白いところ

二つの課題に91チームが参加し、有効な順位表結果704件と技術報告14件を集めた。

どこまで分かった?

要旨はチャレンジ全体を概説しており、個別システムの詳細な性能値や一般化の範囲は示していない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

本論文は、効果的な多言語会話音声言語モデルの開発を進めるために開催された、Interspeech 2026の第2回多言語会話音声言語モデル(MLC-SLM)チャレンジをまとめる。二つの課題である、多言語会話音声での話者分離と認識、および多言語会話音声の理解について説明する。併せて、公開した実際の会話音声データセット、評価手順、基準システムを紹介する。世界各地から91チームが参加し、二つの課題で有効な順位表の結果が704件、技術報告が14件寄せられた。参加システムに基づいて代表的な方法をまとめ、今後の研究に役立つ多言語会話音声の認識と理解に関する実践的な知見を整理する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-23(UTC)
最新改訂
2026-09-23 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

This paper summarizes the Interspeech2026 second Multilingual Conversational Speech Language Model (MLC-SLM) Challenge, which aims to advance the development of effective multilingual conversational speech language models. We describe the two challenge tasks: multilingual conversational speech diarization and recognition, and multilingual conversational speech understanding, together with the released real-world conversational speech dataset, evaluation protocols, and baseline systems. The challenge attracted 91 teams worldwide, with 704 valid leaderboard results and 14 technical reports across the two tasks. Based on the participating systems, we summarize representative approaches and distill practical insights into multilingual conversational speech recognition and understanding to support future research in the community.

arXiv ID: 2609.27514 / 要約の誤りについて