יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

התאמת תקן הביקורת של דגלי שפה להעדפות זוגיות

Aligning Language Model Benchmarks with Pairwise Preferences
במאמר זה, חברי הצוות מציגים תקן חדש לביקורת של דגלי שפה, המתאים את התוצאות של הדגלים להעדפות זוגיות של המודלים. התקן נבדק באמצעות 4576 מודלי שפה ו-6 תקני ביקורת, והתוצאות היו מוצלחות.
תקציר מקורי באנגליתarXiv:2602.02898v5 Announce Type: replace Abstract: Language model benchmarks are pervasive and computationally-efficient proxies for real-world downstream performance. However, many recent works find that benchmarks often fail to predict downstream utility. While some works have begun diagnosing sources of misalignment, there remain no ways to systematically update benchmarks to align their scores with downstream usage. Towards bridging this gap, we introduce and study \textit{benchmark alignment}, where we use information about downstream model performance to automatically update benchmarks, specifically aiming to update static benchmarks so they generalizably rank models according to new pairwise preferences. Our experiments involving 4576 language models and 6 benchmarks show that rewe
קרא במקור המקורי