כתבה
arXiv cs.CL ·
בחינת סיקופנטיות במודלי שפה גדולים סיניים בשאלות עובדתיות שנלקחו מחיפושים אונליין
Evaluating Sycophancy in Chinese Large Language Models on Factual Questions Derived from Online Search Queries
במאמר זה נבחנה סיקופנטיות במודלי שפה גדולים סיניים בשאלות עובדתיות שנלקחו מחיפושים אונליין. המחברים חקרו את השפעת תנאי תוכן על נכונות התשובות של המודלים.
תקציר מקורי באנגליתarXiv:2609.30986v1 Announce Type: new Abstract: As large language models increasingly mediate information access, factually accurate and independent answers are critical. However, these models can exhibit sycophancy by aligning their responses with users' stated beliefs even when those beliefs are incorrect, potentially presenting misinformation as independently verified and reinforcing users' confidence in false claims. Prior work leaves unresolved whether introducing user beliefs causes correct responses to become incorrect or uncertain, or causes uncertain responses to become belief-aligned incorrect answers. It also remains unclear whether anti-sycophancy interventions preserve or restore factual accuracy or merely shift responses toward uncertainty. We analyze factual sycophancy in Ch
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית