כתבה
arXiv cs.CL ·
מעבר מתקשרות אורקל: בדיקת תיאום כוונות אינטראקטיביות תחת תקשורת כושלת וכוונות שונות
Beyond Oracle Communication: Benchmarking Interactive Intent Alignment Under Miscommunication and Evolving User Intent
במאמר זה נבחן תיאום כוונות אינטראקטיביות תחת תקשורת כושלת וכוונות שונות, כולל שימוש ב-GPT-5.
תקציר מקורי באנגליתarXiv:2609.38604v1 Announce Type: new Abstract: Modern LLM agents increasingly tackle complex tasks through interactive, long-horizon exchanges with users, while existing benchmarks generally assume that users always accurately and sufficiently communicate a fixed intent. However, this oracle communication assumption rarely holds in practice: users may miscommunicate, change their goals, and run out of patience. We define this task setting as Interactive Intent Alignment, where agents must recover and continuously track the user's current intent despite imperfect communication and evolving goals. To study this setting, we introduce Drift-Bench++, a principled benchmark construction pipeline for verified executable tasks with controlled misalignment and intent shifts, along with an interact
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית