כתבה
arXiv cs.LG ·
Persistent Depth Ordering amid Shifting Block-Bypass Responses in Language Model Pretraining
תקציר מקורי באנגליתarXiv:2610.01165v1 Announce Type: new Abstract: Layer interventions are widely used to probe the internal organization of language models, yet most analyses examine a single training checkpoint even though model representations and computations evolve throughout pretraining. This leaves open which depth-dependent intervention responses reflect persistent organization and which are transient consequences of training. We study this question using single-block identity bypass on fixed teacher-forced contexts across five released trajectories and 11 model-domain combinations. We find that block-bypass responses retain recognizable depth ordering while their magnitudes redistribute: nearby checkpoints preserve stronger rank correspondence than distant ones, and large changes concentrate at posi
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית