יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

כאשר הגמר הקנוני הוא רע: פורמליזציה ומדידה של הקפיצה במודלי שפה גדולים

When the Canonical Completion Is Wrong: Formalizing and Measuring the Jump in Large Language Models
מחלוקת סביב יכולת מודלי שפה גדולים לבצע קפיצות אבדוקטיות. חוקרים חקרו 14 מודלים ומצאו שחלקם יכולים לבצע קפיצות.
תקציר מקורי באנגליתarXiv:2608.26187v3 Announce Type: replace-cross Abstract: Whether large language models (LLMs) can perform the abductive leap from evidence to a new system of axioms, commonly referred to as a jump, has recently attracted considerable debate. A prominent position holds that LLMs are structurally incapable of such jumps, while recent studies challenge both its mechanism and empirical evidence. One of the main reasons why the debate remains open is the difficulty of defining the jump precisely enough to test it. In this paper, we attempt to develop a formal account of the jump in four steps and measure the second. These steps ask what the default completion of partial data is, when the constraints exclude it, whether the new structure agrees with later observations, and how successive jumps
קרא במקור המקורי