יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

טריינינג מפוצלת מודעת למשימות למודלי שפה גדולים

Task-Aware Federated Fine-Tuning for MoE-based Large Language Models
טריינינג מפוצלת מודעת למשימות למודלי שפה גדולים. FedTAR, טכניקה חדשה, מגדילה את יעילות הטריינינג בסביבות פוצלות.
תקציר מקורי באנגליתarXiv:2609.13395v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become a widely adopted architecture for Large Language Models (LLMs), as it improves model capacity while limiting computational overhead through sparse expert activation. This property makes MoE-based LLMs particularly attractive for resource-constrained distributed environments. However, federated fine-tuning of MoE-based LLMs remains challenging under heterogeneous client data. Since clients often correspond to different task preferences, directly aggregating their local updates may weaken expert specialization and introduce conflicting update directions on shared experts. To address these challenges, we propose FedTAR, a task-aware federated fine-tuning method for MoE-based LLMs. FedTAR establishes the asso
קרא במקור המקורי