יום שישי, 31 ביולי 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

הבחנה במידע במודלים של שפה גדולים

Information Discernment in Large Language Models
מודלים של שפה גדולים מתקשים להבחין בין מקורות מידע מהימנים ופחות מהימנים. מחקר חדש מציג את Learn2Discern, כלי לבחינת יכולת ההבחנה של מודלים אלו.
תקציר מקורי באנגליתarXiv:2607.19355v1 Announce Type: new Abstract: LLMs are increasingly used with external knowledge sources like the internet. Do they weigh information appropriately -- updating more for reliable sources (source discernment) and more when claims bring priors closer to the truth (truth discernment)? We formalize this as information discernment and introduce Learn2Discern (L2D), an experimental framework and benchmark grounded in three normative axioms with interpretable metrics. To establish external validity, a pre-registered, quota-matched user study (n=299) confirms that real LLM users endorse all three axioms and report that violations reduce their trust and usage intent. Across 13 models and nearly 670K trials, we find consistent failures across both dimensions: models perform near cha
קרא במקור המקורי