כתבה
arXiv cs.LG ·
RubricRefine: שיפור יעילות כלי-שימוש עם תיקון-טריינינג חינם
RubricRefine: Improving Tool-Use Agent Reliability with Training-Free Pre-Execution Refinement
RubricRefine משפרת אמינות כלי-שימוש עם תיקון-טריינינג חינם. השיטה יוצרת רוביקים ספציפיים לכלי ובודקת את הקוד כנגדם. רוביקים אלה מסופקים על ידי תיעוד הכלי. השיטה נבדקה על מספר כלים והוכיחה יעילות.
תקציר מקורי באנגליתarXiv:2605.09730v5 Announce Type: replace Abstract: Iterative self-refinement is a popular inference-time reliability technique, but its effectiveness in code-mode tool use depends heavily on the structure of the feedback signal: unstructured critique helps inconsistently across models, and even revision with real execution feedback improves only modestly. The dominant failures are inter-tool contract violations (wrong output shape, incorrect tool routing, broken argument provenance) that run to completion without raising errors, making runtime feedback insufficient. We introduce RubricRefine, a training-free method for pre-execution contract checking that generates task- and registry-specific rubrics, scores candidate code against explicit contract checks, and iteratively repairs failures
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית