יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

SEA-SpeechBench: בנך להבנת דיבור בדרום-מזרח אסיה

SEA-SpeechBench: A Large-Scale Multitask Benchmark for Speech Understanding Across Southeast Asia
SEA-SpeechBench הוא בנך גדול-קנה מידה להבנת דיבור ב-11 שפות בדרום-מזרח אסיה. הבנך כולל 97,194 דגימות ו-597 שעות של נתוני אודיו. הוא בודק 9 משימות שונות, כולל הכרת דיבור, תרגום דיבור וזיהוי רגשות. התוצאות מראות פערים משמעותיים בביצועים בין מודלים שונים.
תקציר מקורי באנגליתarXiv:2609.09672v1 Announce Type: new Abstract: The rapid advancement of audio and multimodal large language models has unlocked transformative speech understanding capabilities, yet evaluation frameworks remain predominantly English-centric, leaving Southeast Asian (SEA) languages critically underrepresented. We introduce SEA-SpeechBench, to the best of our knowledge, the first large-scale multitask benchmark that evaluates speech understanding in 11 SEA languages through 97,194 samples across 99 evaluation sets and 597 hours of curated audio data. Our benchmark comprises 9 diverse tasks across 3 categories: speech processing (automatic speech recognition, speech translation, spoken question answering), paralinguistic analysis (emotion, gender, age, speaker recognition), and temporal unde
קרא במקור המקורי