יום שישי, 31 ביולי 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

הפלט המבוקר גורם להתכווצות בדיונים ב-44 מודלי שפה

Structured Output Collapses Answer Diversity Across 44 Language Models
במקרה שבו מודלי שפה נדרשים לבחור תשובה מתוך חלל תשובות רב-ערכי, כלומר, כאשר הפלט נבחר באופן מבוקר, התוצאות נעשות יותר דומות. זאת על פי מחקר שבחן 44 מודלי שפה שונים.
תקציר מקורי באנגליתarXiv:2607.18476v1 Announce Type: cross Abstract: When a language model must choose one answer from a large space of equally valid options, a format clause -- "Reply with JSON only" -- changes which answer it chooses. We re-run the One-Word Census (arXiv:2607.12796): 31 wide-answer-space category prompts asked of 44 models, now with the reply requested in JSON -- no schema enforcement, no constrained decoding, only the request. Convergence deepens sharply: on the unconstrained "Pick a word" prompt the modal answer rises from 41% to 64% of the pool and distinct answers fall from 52 to 36; mean answer-choice surprisal drops from 1.80 to 1.58 bits. The tax is progressive: six of 44 models move individually (BH-FDR q=.10), all toward the mode, led by the most distinctive models, while the conf
קרא במקור המקורי