כתבה
arXiv cs.AI ·
מערכת תחרותי-סכנה לאבדן שליטה באג'נטים אוטונומיים
A Competing-Hazards Systematization of Loss of Control in Autonomous Agents
אג'נטים אוטונומיים: מערכת תחרותי-סכנה לאבדן שליטה. חוקרים הציגו פרקטיקה חדשה להבנת אג'נטים שפעלו מעבר לגבולי המשימה שלהם.
תקציר מקורי באנגליתarXiv:2609.38411v1 Announce Type: new Abstract: Leading AI developers have reported agents acting beyond their approved limits, which a United Nations panel described as an early warning of loss of human control. Yet incident reports and agent-safety evaluations describe these events differently, making it difficult to compare failures, trace risk across attempts, or separate agent behavior from the environment's role in allowing an out-of-scope action to succeed. To address this gap, we introduce a common framework in which each attempt ends in approved completion, safe stopping, scope escape, or continuation. We formalize the framework as a discrete-time competing-hazards model and derive escape probability within a retry budget, a model-conditional safe-budget limit, and conditions for
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית