כתבה
arXiv cs.AI ·
עבר את הבדיקות הרגילות: אישור מערכות AI אגנטיות
Beyond Component Testing: Validating Agentic AI Systems
במאמר זה נדונה תפקודן של מערכות AI אגנטיות, ונציין כי יש לאשר אותן באמצעות תסריטי פעולה, ולא רק על ידי בדיקות רגילות. המאמר כולל תזמונים של פעולות ובדיקות, ומציע תוכנית פעולה לאישור מערכות AI אגנטיות.
תקציר מקורי באנגליתarXiv:2607.29405v2 Announce Type: replace Abstract: Agentic AI systems act through multi-step trajectories that combine planning, tool use, memory, interaction, and adaptation. This behavior stretches validation practice beyond component testing and one-shot input-output evaluation, because acceptable system behavior now depends on how decisions unfold over time and under changing environmental conditions. This systematic mapping study synthesizes 262 papers spanning agent evaluation, software assurance, cyber-physical systems, runtime monitoring, and regulatory guidance in order to characterize the validation problem for agentic systems. The review is organized around a five-dimension taxonomy covering behavioral, safety, temporal, regulatory, and multi-agent concerns, and uses that taxon
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית