כתבה
arXiv cs.LG ·
משוואת אופטימליות של בלמן לפלסטיות
A Bellman Optimality Equation for Plasticity
במאמר זה, המחברים עוסקים באופטימליזציה של פלסטיות בתהליכי החלטה מרקוב. הם מציגים משוואת אופטימליות של בלמן לפלסטיות, הדומה למשוואת אופטימליות של בלמן לאמפוארמנט.
תקציר מקורי באנגליתarXiv:2609.10776v1 Announce Type: new Abstract: In continual reinforcement learning, carefully managing the stability-plasticity tradeoff remains a core challenge. Recent work by Abel et al. (2025) formalized this dilemma by defining plasticity as the generalized directed information from an agent's observations to its actions, and empowerment as the generalized directed information from its actions to its observations. This formulation successfully reframes the traditional stability-plasticity tradeoff as an empowerment-plasticity tradeoff. However, while extensive literature exists on optimizing for empowerment, there is currently no research addressing the optimization of plasticity under this new definition. This paper presents preliminary work toward optimizing plasticity within Marko
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית