יום חמישי, 30 ביולי 2026 LIVE
AI־INFO

כתבה MIT Tech Review AI ·

לקות בסיסית הופכת מודלים גדולים לפגיעים

A fundamental flaw leaves LLMs strikingly vulnerable to attack
חוקרים גילו לקות בסיסית במודלים גדולים שהופכת אותם לפגיעים להתקפות. הלקות נובעת מאופן שבו המודלים מזהים הוראות. החוקרים הצליחו לגרום למודלים פופולריים לחשוף מידע שלא היה אמור להיחשף.
תקציר מקורי באנגליתIt is impossible to make large language models fully secure against hacks because of a fundamental flaw in how they work, a team of researchers argue in a paper presented at the International Conference on Machine Learning , a top AI conference, this month. The claim has huge implications for the safety of this technology, which is being used in more and more applications, from government and military systems to online shopping and health care . By taking advantage of this flaw, which concerns how LLMs identify who or what is giving them instructions, the researchers were able to make popular LLMs spit out information they had been trained not to provide, such as how to synthesize cocaine and how to sabotage a commercial aircraft’s navigation system. “There’s a real probability that this i
קרא במקור המקורי
technologyreview.com פתח כתבה מקורית