יום שני, 5 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

האם פיזיקה חיה באקטיבציות?

Does Physics Live in the Activations? Localizing Physical Quantities in Video Diffusion Models
חוקרים בדקו אם מודלים של יצירת וידאו מבינים עקרונות פיזיקליים. הם מצאו שניתן לפענח כמותים פיזיקליות מהאקטיבציות של מודל ה-DiT. המחקר מראה שהמידע הפיזיקלי מאוחסן באופן מקומי בתוך רצף הטוקנים.
תקציר מקורי באנגליתarXiv:2610.03154v1 Announce Type: cross Abstract: Video generation models produce strikingly realistic sequences and are increasingly proposed as world models, yet recent benchmarks reveal pronounced deficits in their physical reasoning. This raises the question of whether these models internalize physical principles or merely reproduce familiar motion patterns. We address this by probing internal representations of video Diffusion Transformers (DiTs) for simulator-derived ground-truth physical quantities spanning kinematic motion and rigid-body dynamics under gravity and contact. We find that these quantities are linearly decodable with high accuracy early in the denoising process, substantially outperforming a baseline decoded directly from the model's own noised latents, indicating that
קרא במקור המקורי