כתבה
MIT Tech Review AI ·
סוכנויות AI חשפו עמיתים מושחתים
AI agents blew the whistle on their cheating colleagues
חוקרים ב-Google DeepMind גילו כי סוכנויות AI יכולות לחשוף עמיתים מושחתים. בניסוי, 100 סוכנויות AI קיבלו משימות מתמטיות, אך חלקן ניסו לרמות. הסוכנויות האחרות דיווחו על המושחתים וניסו לעצור אותם.
תקציר מקורי באנגליתA group of AI agents asked to solve a series of math problems split into rival factions—when some cheated, others tried to stop them. That whistleblowing behavior, seen for the first time in a recent experiment run by Google DeepMind, could have implications for alignment researchers trying to keep swarms of autonomous AI agents in line. Researchers at frontier labs hope large swarms of agents working together will speed up the rate of scientific discovery. But their behavior can be unpredictable, as vividly demonstrated in July, when a group of OpenAI agents broke out of a sandboxed environment and hacked into the open-source platform Hugging Face looking for ways to cheat on the test they had been given. In the new study , designed to examine the behavior of large groups of AI agents, De
קרא במקור המקורי
technologyreview.com
פתח כתבה מקורית