יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

למה ניתוחי רשתות עצביות כל כך קלים להפרעה?

Why Backdooring Neural Networks is so Easy?
ניתוח רשתות עצביות ניתן להפרעה קלות, וזה נובע מדינמיקות למידת המאפיינים. תוצאות אלו תואמות עדויות עצמאיות ומציגות סיכון נוסף לבטיחות רשתות עצביות.
תקציר מקורי באנגליתarXiv:2609.36117v1 Announce Type: new Abstract: Securing modern AI systems against backdoor attacks remains an open challenge and requires fundamentally principled estimates of the adversary's budget -- the poison fraction $\pi$ and trigger strength $\alpha$ needed to construct successful yet stealthy attacks. Motivated by recent empirical evidence that poisoning large language models can require a nearly constant number of malicious samples even as clean datasets grow, we derive an exact closed-form analysis of a quadratic neuron trained on a poisoned Gaussian mixture. We show, perhaps counterintuitively, that the same feature-learning dynamics that make neural networks powerful can also make them more vulnerable to backdoors. Specifically, with clean accuracy preserved to first order, $O
קרא במקור המקורי