יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

מה עוזר לפעילות חזרה בדגמי שפה מעגליים?

What Makes Recurrence Effective in Looped Language Models?
במאמר זה, חוקרים בוחנים את יעילות החזרה בדגמי שפה מעגליים. הם מצאו שחזרה יכולה לשפר תפיסה, אך עלולה להפגיע בידע. הם גם חקרו את השפעת תפריט החזרה והמקום שבו הוא נאפשר.
תקציר מקורי באנגליתarXiv:2609.36636v1 Announce Type: cross Abstract: Looped language models (LoopLMs) increase computational depth through parameter sharing, offering a path to scale inference computation without adding parameters. However, it remains unclear when additional recurrence is beneficial and how architectural choices affect its effectiveness. Through controlled experiments, we systematically examine (1) when recurrence helps, (2) where it should be applied, and (3) how its conditioning affects performance. Our evaluation covers inference budgets below, within, and beyond the training horizon under knowledge and reasoning tasks. (1) We find that recurrence can improve reasoning beyond the training horizon while degrading knowledge performance, but harder reasoning instances do not consistently ben
קרא במקור המקורי