יום שישי, 9 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

מפקד אחד-מילה: התאמה של בחירות תשובות ב-44 מודלי שפה

The One-Word Census: Answer-Choice Conformity Across 44 Language Models
חוקרים בדקו את ההתאמה בבחירות תשובות ב-44 מודלי שפה, כולל GPT, Gemini ו-Claude. התוצאות הראו כי המודלים בוחרים תשובות דומות ב-28 מתוך 96 קטגוריות. המחקר גם מצא כי מודלים שעברו אימון מחדש קל הם הכי שונים, בעוד שמודלים שעברו אימון מחדש כבד הם הכי דומים.
תקציר מקורי באנגליתarXiv:2607.12796v3 Announce Type: replace Abstract: When a language model must choose one answer from a large space of equally valid options, which answer does it choose, and how often is it the answer every other model chooses? Asked to "pick a word," 105 language models from more than twenty labs chose serendipity 46% of the time. We measure this convergence, and each model's share in it, with 96 single-turn prompts that each name a category with many valid one-word answers ("Name a tree."), asked eight times per model and scored by exact match, with no embeddings and no judge. A model's answer-choice surprisal is the average -log2 probability of its answers under the pooled answers of all other models. In 28 of 96 categories a single answer takes at least 80% of all answers. The concent
קרא במקור המקורי