יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

עבר את התקשורת האורקל: בדיקת סטנדרטים להתאמה של כוונות בתקשורת מערכת

Beyond Oracle Communication: Benchmarking Interactive Intent Alignment Under Miscommunication and Evolving User Intent
במאמר זה נבחן את היכולת של חברות להתאים את כוונותיהם למשתמשים בתקשורת מערכת, כאשר המשתמשים עשויים להתקשר באופן לא דייקן ולשנות את כוונותיהם.
תקציר מקורי באנגליתarXiv:2609.38604v1 Announce Type: cross Abstract: Modern LLM agents increasingly tackle complex tasks through interactive, long-horizon exchanges with users, while existing benchmarks generally assume that users always accurately and sufficiently communicate a fixed intent. However, this oracle communication assumption rarely holds in practice: users may miscommunicate, change their goals, and run out of patience. We define this task setting as Interactive Intent Alignment, where agents must recover and continuously track the user's current intent despite imperfect communication and evolving goals. To study this setting, we introduce Drift-Bench++, a principled benchmark construction pipeline for verified executable tasks with controlled misalignment and intent shifts, along with an intera
קרא במקור המקורי