כתבה
arXiv cs.CL ·
YANchor-4B: נימוק ארוך טווח יעיל
YANchor-4B: Effective Long-Horizon Reasoning in O(N) Time with O(1) Memory
YANchor-4B הוא מודל רקורנטי כללי השומר זיכרון חיוני כדי לאפשר נימוק ארוך טווח יעיל. המודל משיג תוצאות מרשימות בבעיות מתמטיות ומעלה ביצועים משמעותיים על פני מודלים אחרים.
תקציר מקורי באנגליתarXiv:2610.10118v1 Announce Type: cross Abstract: Long-horizon reasoning demands access to earlier information at a manageable generation cost. Full-history attention incurs growing storage and computation, while recurrent compression can lose precise details. Therefore, we present YANchor-4B, a general-purpose recurrent model that preserves crucial memory as ANchors for retrieval during subsequent reasoning. Beyond $O(N)$-time generation and $O(1)$ memory, YANchor enables effective long-horizon reasoning through its multidimensional memory mechanism. For example, on challenging math problems, it achieves 82.93% mean pass@1 on AIME 2024--2026 and 63.64% on HMMT, substantially outperforming linear-time, constant-state counterparts, including larger models. It also delivers several-fold high
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית