יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

HoneyRoute: רשת-המלכודת למודלי LLM לשירות חסין

HoneyRoute: Honeypot-Model Routing for Adversarial LLM Serving
HoneyRoute מזהה בקשות מפולשות ומפנה אותן למודל-המלכודת. המערכת נבנתה כדי לשגר בקשות פולשניות למודל-המלכודת, ולאפשר ניתוח של התנהגות הפולש. המערכת נבחנה במספר תרחישים, והראתה יעילות גבוהה בזיהוי ובהפניית בקשות פולשניות.
תקציר מקורי באנגליתarXiv:2609.08306v3 Announce Type: replace-cross Abstract: We introduce HoneyRoute, an inference-serving layer that detects whether an incoming request is malicious and, if so, routes it to a dedicated honeypot model, shielding production while the adversary's interaction is continuously harvested for intelligence. Existing defenses embed traps inside model memory or rebuild deception at the protocol layer, leaving the serving tier unprotected and feeding nothing back into detection. HoneyRoute couples (i) a streaming router (a frozen 0.8B-embedding backbone with per-domain MLP heads), (ii) a dual-implementation honeypot (a rule/prompt-engineered code honeypot or a dedicated same-family replica), and (iii) an analysis loop that converts trapped interactions into attacker fingerprints for ro
קרא במקור המקורי