יום שישי, 9 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

עבודה מוקדמת: למידת מודלים על-גדולה - חידוש ניתוח

Poster: A Preliminary Study of LLM Distillation Inference
במאמר זה נחקרה שיטה לזיהוי תקיפות של חברות על-גדולה. השיטה, הנקראת 'חידוש ניתוח', מאפשרת לזהות אם מודל נתון הוא תוצר של תקיפה או שהוא נלמד באופן independent. המאמר כולל תוצאות ניסויים שהראו כי השיטה יעילה.
תקציר מקורי באנגליתarXiv:2610.12137v1 Announce Type: cross Abstract: Unauthorized model distillation, in which a model is trained on the outputs of a proprietary large language model (LLM), is a growing threat to model providers. We study distillation inference: determining whether a suspect model was distilled from another model or trained independently. We formulate this problem as a hypothesis test and estimate the behavior expected under each hypothesis by training shadow models: distilled shadow models learn from the teacher's reasoning traces, whereas independent shadow models learn only from reference answers. The auditor measures how closely each model predicts the teacher's reasoning outputs and then uses the shadow models to convert the suspect's score into a calibrated p-value. In a preliminary st
קרא במקור המקורי