כתבה
arXiv cs.AI ·
Learning from Revision Consequences: Hindsight Meta-Experience Distillation for Self-Improving Agents
תקציר מקורי באנגליתarXiv:2610.07979v1 Announce Type: new Abstract: As agents continuously improve by generating and revising Skills, the process that discovers and refines those Skills becomes a learnable object in its own right. Task-Skills directly act on task execution, whereas Meta-Skills govern how agents discover and improve future Skills; their value therefore emerges through the subsequent search processes they induce. Existing approaches improve Meta-Skills from observed raw Skill-search trajectories and branch outcomes. However, branch performance entangles the effects of the initial discovery state and the Meta-Skill revision that generated the search process, making it difficult to characterize what a particular revision actually changed, and pushing updates toward revisions that benefit from fav
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית