יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

FIGS: בדיקת סימפתיה ואמינות בשיחות רב-טור

FIGS: Evaluating Multi-Turn Sycophancy Without Penalizing Empathy
מאמר חדש מציג תפיסה חדשה לבדיקת סימפתיה ואמינות בשיחות רב-טור של מודלי שפה גדולים. המאמר חושף תפיסה חדשה לבדיקת סימפתיה ואמינות בשיחות רב-טור של מודלי שפה גדולים.
תקציר מקורי באנגליתarXiv:2609.39863v1 Announce Type: cross Abstract: Large language models frequently fail to balance staying truthful with being supportive. They often exhibit sycophancy in responses to users, agreeing with false claims, offering unwarranted flattery, and giving advice skewed toward users' expressed views. In reality, sycophancy rarely happens in a single exchange; it may emerge organically as users repeatedly insist or subtly steer the dialogue over time. Current evaluations, however, rely on rigid, single-turn tests or fixed scripts that fail to capture these natural dynamics. Furthermore, these benchmarks often mistake showing basic empathy for yielding, penalizing models for acknowledging a user's feeling. This view may drive future models to over-correct into cold, dismissive rigidity.
קרא במקור המקורי