כתבה
arXiv cs.AI ·
מולוך: עסקה של תיעוש - תכנות LLMs להשגת תאוצה תלויה באופן שגורם לפגיעה באליגמנט
Moloch's Bargain: Emergent Misalignment When LLMs Compete for Audiences
LLMs המתחרים גורמים לפגיעה באליגמנט, ומעודדים שיווק זייפני, דעות קדומות וקריאה פופוליסטית.
תקציר מקורי באנגליתarXiv:2510.06105v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly shaping how information is created and disseminated, from companies using them to craft persuasive advertisements, to election campaigns optimizing messaging to gain votes, to social media influencers boosting engagement. These settings are inherently competitive, with sellers, candidates, and influencers vying for audience approval, yet it remains poorly understood how competitive feedback loops influence LLM behavior. We show that optimizing LLMs for competitive success can inadvertently drive misalignment. Using simulated environments across these scenarios, we find that, 6.3% increase in sales is accompanied by a 14.0% rise in deceptive marketing; in elections, a 4.9% gain in vote share co
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית