כתבה
arXiv cs.AI ·
למידת מיומנויות מטא לעיצוב חרישי סוכן באי-אי-4-אי
Learning Meta-Skills for Agent Harness Design in Test-Time AI4AI
במאמר זה, חוקרים חוקרים את יכולת הלמידה של סוכן לעיצוב חרישי טוב יותר לסוכן נתון. הם מציגים רעיון של Meta-Skill, שהוא קבוצה של עקרונות שמספרים כאשר סוכן צריך סיוע ומה סוכן צריך לספק. הם מדגימים את Meta-Skill באמצעות ניסויים על Harness-Bench ו-NewtonBench.
תקציר מקורי באנגליתarXiv:2609.38143v2 Announce Type: replace Abstract: Agent performance depends on both reasoning ability and the environment in which it acts. We study test-time AI-for-AI, asking how a Builder can learn to construct better execution environments for a Target while both models' weights remain fixed. To make the Builder's experience reusable, we introduce Meta-Skill: principles specifying when support is needed and what resources to provide. The Builder learns these principles from Target's execution feedback on the development set, then uses the frozen skill bank to construct harnesses for unseen tasks. Across Harness-Bench and NewtonBench, full-bank meta-skills improve macro-average performance by 8.95 percentage points over no-skill construction, and 12.02 points over direct delivery of t
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית