יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

עקביות עצמית אסטרטגית

Strategic Self-Consistency
חוקרים גילו כי ספקים יכולים לנצל את המודלים Llama ו-Qwen על ידי יצירת נתיבי היגיון נוספים. המחקר בדק את השפעת הדבר על מודלים שונים, כולל DeepSeek-R1.
תקציר מקורי באנגליתarXiv:2609.30352v1 Announce Type: new Abstract: Self-consistency has become a popular technique for enhancing the reasoning abilities of large language models by generating multiple reasoning paths and selecting the final answer through a majority vote. However, because model providers typically charge users in proportion to the number of reasoning paths generated, they have a financial incentive to artificially increase the path count. In this work, we show that an unfaithful provider can exploit this incentive using a simple, efficient algorithm while avoiding detection by an auditor: by generating and strategically reordering additional reasoning paths, the algorithm makes every path appear necessary to reach the majority. To validate our algorithm, we conduct experiments with multiple
קרא במקור המקורי