יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

מבעד למראה: קריאה וכתיבה ישירה של טרנספורמרים

Through the Looking Glass: Directly Reading and Writing Transformers
חוקרים גילו דרך לקרוא ולכתוב טרנספורמרים ישירות, מה שמאפשר להבין טוב יותר את התהליכים הפנימיים שלהם. המחקר מראה כי ניתן לזהות את הרכיבים החשובים ביותר בטרנספורמר ולהבין כיצד הם תורמים לתוצאות.
תקציר מקורי באנגליתarXiv:2609.10210v1 Announce Type: new Abstract: How many of a transformer's components decide a token? Counted by the absolute value of each unit's and channel's contribution to the logit, one prediction rests on thousands to hundreds of thousands of them. But contributions are signed, and across eighteen models the mass pushing away from the predicted token is a median of seven times the mass carrying it. Divide by the net and the count is dozens: on the baseline, 53 components carry ninety percent of a prediction, 13 it cannot survive losing, and 8 suffice to produce it alone. Across twelve models trained elsewhere, 124M to 7B parameters, the sufficient set runs from two components to sixteen, and what a prediction draws on, followed all the way back, is one to three percent of the model
קרא במקור המקורי