כתבה
arXiv cs.LG ·
עקביות עצמית אסטרטגית
Strategic Self-Consistency
חוקרים גילו כי ספקים יכולים לנצל את המודלים Llama ו-Qwen על ידי יצירת נתיבי היגיון נוספים. המחקר בדק את השפעת הדבר על מודלים שונים, כולל DeepSeek-R1.
תקציר מקורי באנגליתarXiv:2609.30352v1 Announce Type: new Abstract: Self-consistency has become a popular technique for enhancing the reasoning abilities of large language models by generating multiple reasoning paths and selecting the final answer through a majority vote. However, because model providers typically charge users in proportion to the number of reasoning paths generated, they have a financial incentive to artificially increase the path count. In this work, we show that an unfaithful provider can exploit this incentive using a simple, efficient algorithm while avoiding detection by an auditor: by generating and strategically reordering additional reasoning paths, the algorithm makes every path appear necessary to reach the majority. To validate our algorithm, we conduct experiments with multiple
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית