יום שני, 5 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

הפשטת LLM-כשופט: דגם ניתוחי טרקטבי לפחתת זמן המשקל

Demystifying LLM-as-a-Judge: Analytically Tractable Model for Inference-Time Scaling
במאמר זה, נוצר דגם ניתוחי טרקטבי לפחתת זמן המשקל ב-LLMs. הדגם מבוסס על רגרסיה לינארית Bayesiana עם סיימפלר כבד-משקל, ומוצג כפתרון לבעיית הפחתת זמן המשקל ב-LLMs.
תקציר מקורי באנגליתarXiv:2512.19905v3 Announce Type: replace-cross Abstract: Recent developments in large language models have shown advantages in reallocating a notable share of computational resource from training time to inference time. However, the principles behind inference time scaling are not well understood. In this paper, we introduce an analytically tractable model of inference-time scaling: Bayesian linear regression with a reward-weighted sampler, where the reward is determined from a linear model, modeling LLM-as-a-judge scenario. We study this problem in the high-dimensional regime, where the deterministic equivalents dictate a closed-form expression for the posterior predictive mean and variance. We analyze the generalization error when training data are sampled from a teacher model. We draw
קרא במקור המקורי