יום חמישי, 8 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

תפיסות בטיחותיות לגבי סוכנים עם זמן פעולה מוגבל

AI Safety Considerations for Agents With Limited Time to Act
ניתוח תאורטי של הגברת הבטיחות של סוכנים באזורים עם זמן פעולה מוגבל. נכון לעכשיו, ניתן להוכיח שאף סוכן פרפקטי לא יכול להבטיח התנהגות בטוחה.
תקציר מקורי באנגליתarXiv:2610.10285v1 Announce Type: cross Abstract: In the wake of the increasingly public discussion about AI alignment, recent work has tried to propose specific AI architectures that behave safely. However, the proposed arguments that seemingly demonstrate proved alignment mostly neglect the environment the agent needs to act in. We discuss theoretical bounds for agent-agnostic safety guarantees in environments that can only be partially observed and within which an action is required within limited time. We introduce two realistic scenarios, one with an infinite state space and one with signal mixture. In these scenarios, we prove that even a perfect agent cannot guarantee safe behaviour. It will be argued that for any proof of AI safety or alignment, the environment and associated safe
קרא במקור המקורי