arXiv論文メモ
新着一覧
cs.AI · 査読状況未確認

音声データセットでのクィア当事者の収録状況を調査

Queer inclusion in speech datasets: An audit and taxonomy of practical tensions

Brooklyn Sheppard, Anaelia Ovalle, Adina Williams and Levent Sagun

この論文をやさしく読む

ひとことで言うと

音声データセットにクィア当事者の声がどれだけ含まれるか、収集方法に何が課題かを調べた。

何に役立つ?

音声技術のデータ収集や格差評価を設計する際の参考になる。

この研究の面白いところ

一般的な六つのデータセットと、当事者が関わって作った二つのデータセットを比較する。

どこまで分かった?

六つのデータセットで測定可能な割合は0~1.4%で、頑健な格差測定には不足すると述べる。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

音声データセットにLGBTQIA+、すなわちクィア当事者の声がどの程度含まれるかを調べ、現在の音声技術用データでその声が少ない理由を理解するため、実務上の緊張関係を分類する。多様な六つの音声データセットを監査したところ、測定可能な当事者の割合は話者の0~1.4%と低く、頑健に格差を測るには足りなかった。このコミュニティを事例として、社会的に周縁化された集団から音声データを集める際の課題と緊張を考察する。比較のため、クィア当事者によって、当事者のために、当事者とともに作られた音声科学分野の追加の二つのデータセットも監査した。AI・音声技術研究で一般的なデータ収集の慣習は、周縁化されたコミュニティが参加する方法で重視される価値と衝突し得ることを指摘し、それらの緊張を分類する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

In this paper, we examine speech datasets for their inclusion of LGBTQIA+, or queer, voices and provide a taxonomy of tensions to better understand why there is a lack of such voices in current speech technology datasets. Through an audit of six diverse speech datasets, we find that measurable queer representation is low (0-1.4% of speakers) - insufficient for robust disparity measurement. We take this community as a case study to consider what challenges and tensions are associated with collecting speech data from marginalized communities. For comparison, we audit an additional two datasets from the speech sciences that were created by, for, and with the queer community. We note that many customs in speech dataset collection efforts in AI and speech technology research may conflict with values emphasized in participatory approaches with marginalized communities, and provide a taxonomy describing these tensions.

著者のコメント

Accepted at Interspeech 2026

arXiv ID: 2609.25491 / 要約の誤りについて