כתבה
arXiv cs.CL ·
Entropy Sentinel: Probing Entropy Traces for LLM Monitoring
תקציר מקורי באנגליתarXiv:2601.09001v5 Announce Type: replace Abstract: Deploying LLMs raises two coupled challenges: (1) monitoring---estimating where a model underperforms as traffic drifts---and (2) prioritization---deciding where to intervene to close the largest performance gaps. We explore whether top-$k$ logprobs---cheap, consumer-accessible signals from standard inference---can serve as reliable proxies for domain-level quality of both verifiable and subjective tasks. We summarize each response's output-entropy profile into a compact vector, predict instance quality with a lightweight classifier, and then average predictions to yield a domain-level estimate. On verifiable tasks (ten STEM benchmarks, nine LLMs, exhaustive train/test compositions), estimates often track held-out accuracy, with several m
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית