יום רביעי, 7 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

מרכיבי העדפה: ניתנים לאמירה, ניתנים לאימות, וניתנים להעברה

Verifiable, Articulable, and Tacit Components of Preference
נוסחאות חדשות לעדפות יצירתיות ולאמירתן. נוסחאות אלה כוללות נתוני עדפה גדולים, כלי חישוב, ומודלי למידה.
תקציר מקורי באנגליתarXiv:2610.03025v2 Announce Type: replace Abstract: What makes a short story gripping; a news article newsworthy; or a math proof elegant? These constructs resist articulation or verification; their meaning is at least partially tacit. However, modern AI models are improved primarily via articulated constitutions, rubrics and verifiers (i.e. in RLAIF and RLVR); tacit components of preferences are typically understudied. We introduce a large, labeled preference dataset CreativePreferences, containing 2.8M texts labeled by 317M human preference judgments across 7 creative domains, with 42 benchmark tasks. We model these labels with executable programs, rubric banks and densely trained models (V, A and VAT, respectively). We observe robust articulability gaps, VAT-VA; and verifiability gaps,
קרא במקור המקורי