כתבה
arXiv cs.LG ·
LLM: מדוע הם "מהללים" תשובות
On the Tip of the Tongue: Why LLMs Hallucinate Answers They Can Decode
חוקרים בודקים מדוע מודלי LLM "מהללים" תשובות, אף על פי שהתשובה הנכונה ניתנת לפענוח ממצבים ביניים. הם מצאו כי הבעיה קשורה לאופן בו המודל בוחר את התשובה הסופית.
תקציר מקורי באנגליתarXiv:2603.13911v2 Announce Type: replace-cross Abstract: A language model can give the wrong answer even when the correct answer is decodable from its intermediate states. To study this gap between decodability and selection, we distinguish \textit{read} from \textit{write} at the first answer token. Read asks whether the gold token can be decoded from intermediate residual states under same-relation decoy controls. Write asks whether the final readout ranks that token first among content tokens. Under three different readers, with a randomized-label control, a substantial fraction of failures remain readable while another content token is selected. We explain this through the selection margin at the final readout, the difference between the answer logit and the logit of its strongest alt
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית