כתבה
arXiv cs.CL ·
4-Tensor Attention Model for Semantic Physical Reality
תקציר מקורי באנגליתarXiv:2610.11716v1 Announce Type: cross Abstract: We describe a 4-tensor attention model that predicts the next semantic state of a scene, for video generation and robot planning. A window of states has positions (x, t) and two fibers, a semantic fiber and a temporal-context fiber, and one softmax normalizes attention jointly over the window. Frames and an agent's situation are written as those states; the encoder, the renderer, and the planner remain outside the update. To test the update on its own, we train on ROCStories, where each window poses the same next-sentence task at the semantic layer. On the validation split, with one seed per setting, the last-sentence cross-entropy on the three matched settings is lower for the 4-tensor model than for a free-running one-dimensional transfor
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית