כתבה
arXiv cs.AI ·
הגדרת סוכנים AI
Defining AI Agents: A Compendium of Criteria, Metrics, and Benchmarks
חוקרים פרסמו סקירה על סוכנים AI, המציעה מסגרת אחידה להערכת יכולות סוכנים. הסקירה מארגנת את השיטות להערכת סוכנים לפי חמישה ממדים: אינטראקציה עם הסביבה, למידה והסתגלות, אוטונומיה, התנהגות מכוונת יעד, ותוהו זמני. החוקרים גם מציגים את 'Agent Compendium', משאב דיגיטלי ציבורי המארגן ומרחיב את שיטות ההערכה.
תקציר מקורי באנגליתarXiv:2609.11018v1 Announce Type: new Abstract: The term agent in artificial intelligence lacks a standard definition, complicating the evaluation, comparison, and reproducibility of AI agent research. We address this ambiguity through a survey organized around five dimensions of agenticness: environmental interaction, learning and adaptation, autonomy, goal-directed behavior, and temporal coherence. For each dimension, we examine how the underlying capability has been conceptualized across prior work and synthesize the metrics, benchmarks, and evaluation frameworks used to assess it. This review provides a structured account of the current landscape of agent evaluation, highlighting both established approaches and areas where evaluation remains limited or inconsistent. We additionally int
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית