כתבה
arXiv cs.CL ·
MoEGen: Mixture-of-Experts for Instance-Adaptive LoRA Generation
תקציר מקורי באנגליתarXiv:2608.03275v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning (PEFT) enables efficient adaptation of large language models, but existing MoE-based PEFT methods typically improve capacity by storing multiple full LoRA experts, causing adapter storage to grow linearly with the number of experts and restricting adaptation to a fixed expert pool. We ask whether MoE-based PEFT can produce instance-specific adaptations without explicitly storing a separate LoRA module for each expert. To address this gap, we propose MoEGen, an adaptation framework that shifts MoE-based PEFT from expert selection to expert-conditioned parameter generation. Instead of storing each expert as a full LoRA adapter, MoEGen represents each expert as a small learnable vector, termed an expert code.
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית