יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

PADM'E: שיטה לסינתזה של נתונים לבדיקת מודלים

PADM\'E: Preference Alignment Data Synthesis for Meta-Evaluation of LM Agent Evaluators
PADM'E היא שיטה לסינתזה של נתונים לבדיקת מודלים של LM. היא מאפשרת לבדוק את ההתאמה בין החלטות של מודלים לשיפוט אנושי. PADM'E משתמשת במודלים קטנים ואינה דורשת מעורבות אנושית. השיטה הראתה שיפור בהסכמה עם שיפוט אנושי.
תקציר מקורי באנגליתarXiv:2609.36086v1 Announce Type: new Abstract: Language models are frequently employed to evaluate other language models. An LM evaluator scoring agentic behaviors across multiple criteria is valuable, provided that its decisions align with human judgment. We call the problem of evaluating this alignment Meta-Evaluation. Tackling it directly is difficult: collecting human data is expensive, absolute scoring is hard to align, and using an LM meta-evaluator recurses the question of trustworthiness. We adopt a reformulation of meta-evaluation as a preference judgment problem: rather than comparing human and LM evaluator scores of a trajectory, we ask whether their implied preferences align. Building on this, we introduce PADM\'E, a data synthesis method that generates reliable criterion-base
קרא במקור המקורי