יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

הטוקן הראשון אינו הכרעה

The First Token Is Not the Verdict: Hidden Costs of Reading LLM Judges Without Generating
חוקרים מצאו כי קריאת פסק דין של LLM מהטוקן הראשון יכולה להיות מוטעית. המנגנון גורם להטיה באחוזים, במיוחד כאשר השופט אינו מוביל עם טוקן פסק דין. החוקרים ממליצים לדווח על שיעור הובלת טוקן פסק דין.
תקציר מקורי באנגליתarXiv:2610.00054v1 Announce Type: cross Abstract: Reading an LLM judge's verdict from the logits of its first generated token is cheap, requires no generation, and is exactly what constrained decoding and likelihood-scoring evaluation harnesses produce. We show that this readout distorts position bias in one direction: it overstates it in every condition we test, so figures obtained this way behave as upper bounds. The mechanism is that judges do not always lead with a verdict token, on 12% to 49% of pairs for three Qwen3 judges and under 3% for Llama-3.1-8B and Phi-3.5-mini, and forcing a read on those pairs returns whichever response was shown first rather than a judgment. Pooled over the 924 pairs where a judge did not commit, the forced read flips on 89.7% of them when the responses ar
קרא במקור המקורי