יום שישי, 9 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

כיצד קובעים כמה נתונים סינתטיים צריכים פרוב פעילות?

Coverage, Not Difficulty, Sets How Much Synthetic Data an Activation Probe Needs
פרובי פעילות צריכים נתונים סינתטיים כדי לסקור מודלי שפה שהוקמו.
תקציר מקורי באנגליתarXiv:2610.10594v1 Announce Type: new Abstract: Activation probes that monitor deployed language models are trained on synthetic conversations, and how many a probe needs is open. We trace learning curves over 10-590 synthetic samples for three monitoring concepts, high-stakes situations, replies harmful to a person, and replies that do not follow the user's instruction, on fourteen held-out evaluation distributions and four probe models, varying the generator LLM and the prompt's detail. The need is set by what is monitored: probes for high-stakes and harmful are within a few hundredths of their plateau from 80 samples on Gemma-3-27B-IT, instruction probes need several times as many, and the ordering holds on three smaller probe models and on real samples (from dev set). Prior work advise
קרא במקור המקורי