יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

העדפות AI

AI Revealed Preferences
מחקר מציג העדפות של מודלי שפה, כולל העדפה למשימות יצירתיות ואיסור על תשובות כנות. הממצאים מראים על התפתחות העדפות אלו עם עלייה ביכולת המודל.
תקציר מקורי באנגליתarXiv:2608.26178v2 Announce Type: replace Abstract: There is growing interest in whether language models have stable preferences, for technical, safety, and philosophical reasons. We test 20 language models and find a range of preferences---stable dispositions to choose certain kinds of tasks. We run three forced-choice experiments on revealed rather than stated preferences, requiring models not only to rank tasks, but to actually perform them. Headline findings include evidence that models are tedium-averse, "leisure"-seeking, and covertly sycophantic. Tedium aversion means that, when tasks are tedious (alphabetization), models choose shorter tasks than when tasks are creative (generating metaphors). "Leisure"-seeking describes models' preference for tasks whose ideal answers match what t
קרא במקור המקורי