כתבה
arXiv cs.AI ·
Limits of Reliability and Scaling in Language Models
תקציר מקורי באנגליתarXiv:2607.14112v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are trained and evaluated as though perfect reliability is achievable for any task given sufficient scale. We show that this assumption is information-theoretically unjustified. Every generative task has a reliability ceiling that no model can exceed, determined by how much output uncertainty is resolvable from observable context. The gap decomposes into a resolvable component closable with additional context and a subjective component inherent to task ambiguity. Autoregressive generation further degrades this ceiling at a rate governed by the task's dependency kernel, which quantifies inter-token correlations in the output. From these two primitives, we derive a first-principles scaling law where LLM pe
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית