יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

RAIM: רובוסטית של גיבוש של דגמי מודל זולים לזיהוי הלוסינציה

RAIM: Robust Aggregation of Inexpensive Models for Hallucination Detection
במאמר זה, החוקרים חקרו אפשרות לשימוש בדגמי מודל זולים לזיהוי הלוסינציה. הם הציגו את RAIM, שיטת גיבוש שמאפשרת לשלב דגמי מודל זולים ליצירת דגם יעיל יותר. המחברים השוו את RAIM לדגם המוביל Claude ומצאו שהוא יעיל יותר בכ-64% מהעלות של הדגם המוביל.
תקציר מקורי באנגליתarXiv:2609.39229v1 Announce Type: new Abstract: Automatic evaluation of faithfulness increasingly relies on a large language model acting as a judge, yet the most reliable judges are proprietary frontier models, costly and ill-suited to high-throughput monitoring. We investigate whether a panel of cheap open-weight judges (4--9B) can be aggregated to stand in for a frontier one, what the substitution sacrifices, and when it is worth making. We propose RAIM, an aggregation scheme robust to the members' correlated errors, coupling a cross-fitted stacked logistic regression with an admissibility test that, read from the members' own outputs, identifies when aggregating them improves on their best member and stays within reach of the frontier judge. We instantiate RAIM with ten judges from dis
קרא במקור המקורי