יום שישי, 9 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

פורנזיקה של שחקני LLM: זיהוי סוגי ה-LLM

Black-Box Forensics for Conversational LLM Agents
במאמר זה, המחברים פיתחו שיטה לזיהוי סוגי LLM של שחקני חברתי, כדי לחשוף סקאמים של AI. השיטה מאפשרת לזהות את ה-LLM שמאחורי שחקן חברתי, ולזהות קשרים בין שחקנים שונים.
תקציר מקורי באנגליתarXiv:2606.22698v2 Announce Type: replace-cross Abstract: As LLM-powered scams proliferate, black-box forensics for conversational LLM agents offers a path to accountability for systems hidden behind anonymous endpoints. Identifying the base model behind a chatbot endpoint (attribution), without model parameter access or knowledge of the hidden system prompt, would let investigators trace AI-enabled scams back to the providers whose models power them. Detecting when two endpoints run the exact same system prompt (fingerprinting), even one novel and unseen, would link individual scams into criminal networks and expose silent API changes. We conduct an empirical investigation of both capabilities. Our attribution classifiers identify the base model behind an agent with 98% accuracy from a fe
קרא במקור המקורי