כתבה
arXiv cs.LG ·
LP-SFT: תפיסה-שומרת-מקומי-לסופר-הטמעה-מולטימודלית-באמצעות-סטרוקטור-אנטרופיה
LP-SFT: Local-Preserving Supervised Fine-Tuning via Multimodal Entropy Structure
LP-SFT היא תפיסה-שומרת-מקומי-לסופר-הטמעה-מולטימודלית-באמצעות-סטרוקטור-אנטרופיה. היא עוזרת לשמור את היכולות הקיימות של דגמי לשון קיימים בעת הטמעה.
תקציר מקורי באנגליתarXiv:2607.04733v2 Announce Type: replace-cross Abstract: Supervised fine-tuning (SFT) is the standard approach for adapting pretrained language models to downstream domains, yet it often improves target-domain behavior at the cost of degrading pre-existing capabilities. Standard cross-entropy fine-tuning promotes only the observed label token and leaves unconstrained how probability mass is redistributed over other plausible alternatives, potentially distorting the rich local preference structure learned during pretraining. We first analyze next-token predictions using Shannon and Renyi entropies, revealing that pretrained models exhibit a regular multimodal entropy structure. These entropy peaks correspond to varying numbers of plausible alternatives, indicating that the base model intri
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית