יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

MP-Bench: בדיקת סוכני קול כמשתתפים בשיחה רב-צדדית

MP-Bench: Evaluating Voice Agents as a Multiparty Conversation Participant
במאמר זה פותחה בדיקה חדשה לבדיקת סוכני קול בשיחות רב-צדדיות. הבדיקה, MP-Bench, בודקת את היכולת של הסוכנים להשתתף בשיחות רב-צדדיות ולהבין את הקונטקסט. התוצאות הראו שהסוכנים עדיין נמצאים בשלבים מוקדמים בפיתוחם.
תקציר מקורי באנגליתarXiv:2609.13076v1 Announce Type: cross Abstract: Conversational voice agents have advanced significantly, offering increasingly natural human-machine interactions through both cascaded and end-to-end architectures. However, while recent benchmarks extensively evaluate dyadic interactions and passive audio comprehension, they largely overlook a prevalent real-world scenario: multi-party conversations. Evaluating agents in these settings is fundamentally more challenging than in dyadic interactions due to the exponentially greater conversational complexity. For voice agents to integrate seamlessly into human group dynamics, they must not only generate contextually appropriate responses but also demonstrate a nuanced understanding of open turn-taking. To address this gap, we introduce Multip
קרא במקור המקורי