כתבה
arXiv cs.LG ·
הערכת פעולות נגדיות
Counterfactual Action Evaluation, Observation Bottlenecks, and Representation Geometry in Joint-Embedding Predictive World Models
נושא המחקר הוא הערכת פעולות נגדיות במודלים עולמיים. החוקרים מציגים פרוטוקול הערכה חדש שבודק את ההבדלים בין תוצאות פעולות שונות. התוצאות מראות כי שגיאת חיזוי נמוכה לא בהכרח מעידה על יכולת המודל להבחין בין תוצאות פעולות.
תקציר מקורי באנגליתarXiv:2610.02860v1 Announce Type: new Abstract: Low latent prediction error does not establish that a world model distinguishes the consequences of its actions. We introduce an evaluation protocol that traces the same intervention through simulator state, raster observations, target embeddings, and predictor outputs. Exact simulator-state forks in a controlled deformable-physics testbed reveal distinct bottlenecks. Changed commands alter particle motion, yet 41.5% of one-step raster pairs are identical. Observation loss is not the whole explanation: among 579 high-visibility counterfactuals, median predictor-to-target response is 0.0051 and 0.0217 across two seeds, falling to 0.0027 and 0.0116 after variance normalization. An isotropic state perturbation matched to the target counterfactua
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית