כתבה
arXiv cs.AI ·
EGGROLL: שיפור אסטרטגיות התפתחות
EGGROLL, Unrolled: Understanding and Improving Low-Rank Evolution Strategies at Scale
EGGROLL משפר את האסטרטגיות להתפתחות של מודלים גדולים. הוא מחליף פרטורבציות גאוסיות צפופות במוצרים גאוסיות בדרגה נמוכה. החידוש מאפשר שיפור בביצועים ובדיוק.
תקציר מקורי באנגליתarXiv:2609.10980v1 Announce Type: cross Abstract: EGGROLL makes evolution strategies (ES) practical for LLMs by replacing dense Gaussian weight perturbations with low-rank Gaussian products, often of rank one. This choice is computationally attractive but geometrically severe: each rank-one perturbation lies in a zero-volume subset of the ambient matrix space, despite having identity covariance. We characterize the mean EGGROLL update field at finite rank and nonzero perturbation radii, then analyze the error of its finite-population estimator. The population field is obtained by applying an explicit resolvent to the gradient of the objective smoothed by the perturbations. We show that the resolvent can introduce a nonconservative component and can reverse the local stability of an optimum
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית