כתבה
arXiv cs.LG ·
When Forgetting is not Catastrophic: On the Mechanics of Spurious Forgetting
תקציר מקורי באנגליתarXiv:2610.08718v1 Announce Type: cross Abstract: Knowledge that a language model appears to forget during finetuning often remains stored and can be recovered, a phenomenon called spurious forgetting. Finetuning on new facts can even produce forgetting that undoes itself: recall of the old facts collapses, recovers as training continues on new facts alone, and only then erodes for good. We seek to understand when such forgetting is not catastrophic. A minimal associative memory reproduces these dynamics with three ingredients: keys with shared structure, concentrated new values, and normalization in the network. Finetuning moves all old representations along a common direction, hiding the old facts while preserving their relative geometry; normalization withdraws this shift once the new f
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית