יום שני, 5 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

מדוע קל לבצע Backdooring ברשתות עצביות

Why Backdooring Neural Networks is so Easy?
חוקרים גילו שרשתות עצביות מודרניות רגישות להתקפות Backdooring. המחקר מראה שהיכולת ללמוד תכונות עשויה להפוך אותן לפגיעות יותר. התוצאות מספקות מנגנון תאורטי התומך בתצפיות אמפיריות.
תקציר מקורי באנגליתarXiv:2609.36117v2 Announce Type: replace Abstract: Securing modern AI systems against backdoor attacks remains an open challenge and requires fundamentally principled estimates of the adversary's budget -- the poison fraction $\pi$ and trigger strength $\alpha$ needed to construct successful yet stealthy attacks. Motivated by recent empirical evidence that poisoning large language models can require a nearly constant number of malicious samples even as clean datasets grow, we derive an exact closed-form analysis of a quadratic neuron trained on a poisoned Gaussian mixture. We show, perhaps counterintuitively, that the same feature-learning dynamics that make neural networks powerful can also make them more vulnerable to backdoors. Specifically, with clean accuracy preserved to first order
קרא במקור המקורי