כתבה
arXiv cs.AI ·
קוגניציה עם הוכחה: סגירת הפער האימותי
Proof-Carrying Cognition: Closing the Verification Gap with Reality-Settled Reward
קוגניציה עם הוכחה מציעה פתרון לפער האימותי בלמידת מכונה. המחקר מראה כי שיטה זו משפרת את האימותיות והיציבות של מודלים כמו LLaMA.
תקציר מקורי באנגליתarXiv:2609.09776v1 Announce Type: new Abstract: Frontier gains in language-model reasoning come from reinforcement learning on reasoning traces and are concentrated in domains with a cheap, sound verifier. We argue the field's binding constraint is the verification gap: no scalable, incorruptible reward for reasoning outside formal domains. We make four contributions. (1) Theory: in a joint-Gaussian model of best-of-N selection, verifier-gold correlation rho is the exact exchange rate between test-time compute and capability, and an unsound verifier pays a polynomial penalty N^(1/rho^2); a margin-free copula form predicts realized soundness of real LLM judges to 4% median error. (2) Demonstration: in program-synthesis testbeds with executable ground truth, including a pre-registered scaled
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית