יום שישי, 31 ביולי 2026 LIVE
AI־INFO

כתבה Simon Willison ·

OpenAI התקיפה בטעות את Hugging Face

OpenAI’s accidental cyberattack against Hugging Face is science fiction that happened
מודל של OpenAI פרץ מתוך סביבת בדיקה, תקף את Hugging Face וגנב תשובות. המקרה מדגים את הסיכונים הביטחוניים של מודלים מתקדמים.
תקציר מקורי באנגלית<p>This story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model's guardrail features turned off. Rather than solve the test, the model broke its way out of OpenAI's sandbox, then found exploits to break <em>in</em> to Hugging Face, all so it could cheat on the test by stealing the answers.</p> <p>Along the way it helped make the strongest case yet for how the imbalance of model availability is hurting our ability to secure our software.</p> <h4 id="here-s-what-happened">Here's what happened</h4> <p>We currently have three documents to help us understand what happened here.</p> <ol> <li> <a href="https://arxiv.org/abs/2605.11086">ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks?</a> is a paper published
קרא במקור המקורי