כתבה
arXiv cs.LG ·
Safe Meta-Reinforcement Learning via Information Space Reachability
תקציר מקורי באנגליתarXiv:2609.15915v1 Announce Type: new Abstract: Meta-reinforcement learning (meta-RL) enables agents to adapt to unseen tasks with limited experience. Despite its promise, the application of meta-RL in real-world tasks is hindered by safety requirements, which have been underexplored in prior work. In this paper, we propose a safe meta-RL framework that explicitly accounts for safety during adaptation. Our key insight is to reason about safety in the information space, which captures both the physical state and the agent's belief over the underlying task. Within this space, we introduce a safety value function that measures the probability of the agent avoiding unsafe regions indefinitely. We show that this function satisfies a self-consistency condition and a Bellman equation, which make
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית