יום חמישי, 8 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

ריכוז מסה קשב לפיענוח דליל

Attention-Mass Condensation for Sparse Decoding
חוקרים מציגים שיטה חדשה לפיענוח דליל באמצעות ריכוז מסה קשב. השיטה נבדקת על מודל Qwen2-0.5B ומראה תוצאים מבטיחים.
תקציר מקורי באנגליתarXiv:2602.06317v3 Announce Type: replace-cross Abstract: Attention-mass concentration creates an opportunity for sparse decoding, but retained mass alone does not guarantee a stable greedy decision: retrieval error, omitted value directions, and recursive decoding all matter. We formalize this distinction with an exact omitted-mass identity and a sufficient downstream margin condition, then characterize a query-dependent mean-pooled block selector. On Qwen2-0.5B, a paired fresh-selection sweep covers supports of 97--769 positions, contexts of 2K--16K, and five prefixes per context. The primary exact-match result is that none of 60 runs remains identical to dense decoding through 128 tokens. Distributional quality is distinct: for supports of at least 193, seven of nine context-support con
קרא במקור המקורי