יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

התאמה כללית של גורם: פרקטיקה רפורמלית לשיפור פוליציה רציפה ושיפור עצמי רציף

Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement
מאמר זה מציג פרקטיקה רפורמלית לשיפור פוליציה רציפה ושיפור עצמי רציף. המאמר דן בקונספט של שיפור עצמי רציף ומציע פרקטיקה רפורמלית לשיפור פוליציה רציפה.
תקציר מקורי באנגליתarXiv:2609.13406v1 Announce Type: cross Abstract: When we speak of recursive self-improvement (RSI), are we speaking of a phenomenon, a mechanism, or a prospect? Towards autonomous and evolving intelligence, RSI is being claimed at many scales, while no single framework that formally describes these emerging instances exists. Its counterpart in the classical realm, iterative policy improvement, is characterized by generalized policy iteration (GPI), a framework of broad applicability with well-understood theoretical properties, but only where the update principle and the evaluation base lie outside the agent. In this paper, we propose Generalized Agent Iteration (GAI), a formal framework that describes iterative policy improvement and RSI as two cases of a single learning paradigm. GAI def
קרא במקור המקורי