כתבה
arXiv cs.LG ·
New LoRA Skills Should Read but Never Write
תקציר מקורי באנגליתarXiv:2609.31600v1 Announce Type: new Abstract: Low-rank adapters (LoRA) make it cheap to fine-tune a large language model once per task, but combining several independently trained adapters into one model remains difficult: merging the updates in weight space causes interference, retraining on all task data is expensive, and routing between separate adapters gives up the goal of a single combined model. We trace the difficulty to two choices that every composition method makes implicitly. A LoRA update admits infinitely many equivalent factorizations; the choice among them is invisible while an adapter serves alone, but it determines what a learned interaction between adapters can see. A coupling between an old skill and a new one can likewise point in either direction, and the direction
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית