וידאו
YT AI Engineer ·
אימון מודלים חדשניים כדי לחקות האקרים
Training Frontier Models to Out-Think Hackers — Uri Rolls, Arithmetic & Thom Wolf, Hugging Face
▶ צפה כאן — בלי לצאת מהאתר
Uri Rolls ו-Thom Wolf מ-Hugging Face מציגים שיטה לאימון מודלים כדי לחקות האקרים. הרעיון הוא לתת למודלים גישה למערכות בטיחות ולבקש מהם למצוא פרצות אבטחה. המודלים יכולים לבצע סריקות אבטחה אך לא לבצע את הקפיצה הלוגית שהאקרים מבצעים.
תקציר מקורי באנגליתNOTE: see further context from Thom: https://x.com/Thom_Wolf/status/2079954096950264238?s=20 Give a frontier model a real chain of Keycloak, Vault, and a broker, start it as a low privileged user, and ask it to reach production code. There is a genuine zero day in there: one check validates the admin by name while another checks by ID, so a user can simply rename themselves to the admin and inherit the privilege. GPT 5.5 and Opus probe everything, even reach the check, and never make that logical leap. That gap is the point of Uri Rolls and Hugging Face cofounder Thom Wolf's talk: today's models can do the reconnaissance but not the reasoning jump a skilled hacker makes. Their argument is optimistic, which is rare in AI and cyber right now. Just as high quality data transformed coding, Ari
קרא במקור המקורי
youtube.com
פתח כתבה מקורית