יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

הבנת הזיות במודלים גדולים של שפה

When the Wrong Key Wins: Understanding and Detecting Hallucinations in LLMs
חוקרים פיתחו שיטוד לזיהוי הזיות במודלים גדולים של שפה. השיטוד מבוססת על פרטורבציה של מילים מפתח ומדידת השפעתן על הניבוי. החוקרים בדקו את השיטוד על מודלים שונים וקבלו תוצאות טובות.
תקציר מקורי באנגליתarXiv:2609.15106v3 Announce Type: replace Abstract: Large language models can hallucinate even when the knowledge required for a correct answer is already available. We study this failure through a latent-key view of inference, where answer selection depends on competition among associations acquired during pretraining. We show that model predictions can be highly sensitive to individual query keywords, that these influential keywords exhibit entity-specific binding, and that their effects are systematically shaped by pretraining frequency. Multiple bindings can also compete and exhibit higher-order interactions within the same query. Based on this mechanism, we introduce a two-stage keyword-perturbation method for hallucination detection. By removing influential keywords and measuring how
קרא במקור המקורי