כתבה
arXiv cs.CL ·
מבנה להוספה: תווי-זיהוי עריכתיים למודלי שפה גדולים
From Construction to Injection: Edit-Based Fingerprints for Large Language Models
מודלי שפה גדולים: פירוש תווי-זיהוי עריכתיים למניעת הפצה לא מורשת.
תקציר מקורי באנגליתarXiv:2509.03122v5 Announce Type: replace Abstract: Reliable model fingerprints are essential for protecting large language models (LLMs) against unauthorized redistribution and commercial misuse. In black-box deployment, verification is hindered by defensive filtering of suspected fingerprint queries, as well as by downstream model modifications that may weaken embedded ownership evidence. These risks require fingerprints to be robust in both construction and injection. For construction, prior paradigms face an imperceptibility trade-off: natural-language fingerprints may be accidentally activated, whereas garbled fingerprints are statistically exposed and easier to filter. For injection, existing methods struggle to preserve persistent trigger--target behaviors under model modification.
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית