יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

Decision Titan: אימון בזמן בדיקה

Decision Titan: Test-Time Training for Long-Term Memory in Offline Reinforcement Learning
Decision Titan הוא מודל המשלב Test-Time Training עם מעבדת מחלטות, לטיפול בתלות ארוכת טווח בלמידת חיזוק. המודל מצליח ללמוד תלות ארוכת טווח עם טווח 20 פעמים ארוך יותר מחלון ההקשר.
תקציר מקורי באנגליתarXiv:2610.01513v1 Announce Type: new Abstract: Long-term dependencies remain a major challenge for sequential decision-making in the field of AI: RNNs suffer from vanishing gradients and the limited expressivity of vector-based hidden states, whilst Transformer-based models are limited by the quadratic scaling of attention. Recent work has proposed tackling this problem with the Test-Time Training (TTT) framework, which stores episodic memories in the parameters of a neural network through gradient descent at both train and test-time. This approach has seen success in the domain of Natural Language Processing, however, to the best of our knowledge it has not yet been applied to the domain of Reinforcement Learning (RL), nor has there been a study analysing how this memory practically func
קרא במקור המקורי