כתבה
arXiv cs.CL ·
סדר יום רב-קהילתי ל- AI לדיבור
A Cross Community Agenda for Speech AI
AI לדיבור משלב טכנולוגיות NLP ו- HCI. המחקר מציע פתרונות לבעיות ב- AI לדיבור, כולל הגדרת מודלים מלאים יותר לתקשורת וזהות. הוא דן ב- AAC ובמשתמשים אחרים שאינם מיוצגים היטב.
תקציר מקורי באנגליתarXiv:2609.13168v1 Announce Type: cross Abstract: Speech AI, any AI system that recognizes, transforms, or generates speech, is built and evaluated across two communities with only a small overlap: technical natural language processing (NLP) venues (e.g., ACL, ICASSP, Interspeech), and sociotechnical HCI venues (e.g., ASSETS, CHI, FAccT). In this position paper, we work toward a cross-community synthesis, organizing our critique around three problems: speech AI operates with an incomplete model of communication; it operates with an incomplete model of identity; and its metrics measure the wrong constructs. We draw on AAC as a setting where these failures are most visible and their stakes highest, alongside other underserved speakers - people who stutter, multilingual speakers, and non-bina
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית