יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

אחזור הוראות בזמן הסקה למודלי שפה קטנים

Instruction Retrieval at Inference Time for Small Language Models
חוקרים מציגים שיטה חדשה לשיפור ביצועי מודלי שפה קטנים. השיטה, הנקראת 'אחזור הוראות', מאפשרת למודלים קטנים לגשת להוראות מותאמות שנכתבו על ידי מודל מורה. ההוראות אלו מספקות רקע, פרוצדורה ואזהרות מפני טעויות נפוצות. השיטה הוכחה כיעילה בתחומי רפואה, משפטים ומתמטיקה.
תקציר מקורי באנגליתarXiv:2510.13935v3 Announce Type: replace Abstract: The facts a language model stores are tied to its parameter count, so small models that fit on edge devices fail on expert problems, which need specialized knowledge and follow multi-step procedures. Fine-tuning for a specific domain or task writes the knowledge into the parameters but must be repeated for every model and domain, and a retrieved passage leaves the model to find the relevant fact and apply it on its own. We introduce instruction retrieval, which distills a teacher model's expertise into a corpus of instructions tailored so that a small model can follow. For each cluster of a domain's problems, the teacher writes one instruction with the background knowledge the cluster depends on, a procedure for that kind of problem, and
קרא במקור המקורי