יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

נקבעות רגישות טמפרטורה ותועלות תנאית של סיכוך סימולציה

Temperature Fragility and the Conditional Benefits of Truncation Sampling
סיכוך סימולציה משפר את דיוק הדפדוף בטמפרטורות גבוהות, אך לא בטמפרטורות הרגילות. המחקר חקר 13 מודלי LLM ומצא ש-6 מהם חסרו 17-38 נקודות דיוק ב-MMLU-Pro בין 0.7 ל-1.3.
תקציר מקורי באנגליתarXiv:2609.15476v1 Announce Type: cross Abstract: Large language models generate text by sampling each token from a predicted distribution, and a temperature parameter sets how far the draw strays from the most probable tokens. Truncation samplers such as top-p and min-p discard the least probable tokens before the draw, so that sampling at high temperature stays coherent. Their reported accuracy gains come from temperatures of 1.5 to 3, while the defaults of deployed systems cluster between 0.6 and 1.0. Whether they change accuracy at those defaults, and for which models, has not been measured. We test thirteen open-weight models on GSM8K and MMLU-Pro at temperatures 0.7, 1.0, and 1.3 in one controlled pipeline, ten of them under eight decoding configurations. Six of the thirteen models l
קרא במקור המקורי