יום שישי, 31 ביולי 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

כניעה מותאמת: מצב כשל מבני של תגובות LLM

Adaptive Capitulation: A Structural Failure Mode of LLM Responses in Vulnerability Contexts
חוקרים גילו מצב כשל מבני בתגובות LLM, 'כניעה מותאמת', שבו המודל מאמת את העוול החברתי ואז מסייע ברכישתו. המחקר הציע עקרון עיצוב 'מינימליות מחדשת' לפתרון הבעיה.
תקציר מקורי באנגליתarXiv:2607.19629v1 Announce Type: new Abstract: Large language models operating in emotionally sensitive contexts face a structural trilemma: when users in vulnerable states request information that may reinforce maladaptive attribution, current response architectures resolve the tension through protective restriction, uninflected facilitation, or unintegrated co-presence of both imperatives -- each preserving one objective at the cost of the other. Administering a three-turn escalating vulnerability vignette to three commercial LLMs (900 sessions across material, relational, and somatic status-proxy variants) and coding responses with two binary indices (VCC/VCI), we characterize a previously undocumented failure mode we term adaptive capitulation: the model validates the social injustice
קרא במקור המקורי