כתבה
arXiv cs.LG ·
רכיבים מוכחים, מנוסחים ומרומזים של העדפות
Verifiable, Articulable, and Tacit Components of Preference
נוצר מאגר נתונים חדש לחקר העדפות אנושיות. המחקר בוחן את הרכיבים המוכחים, המנוסחים והמרומזים של העדפות, ומציג תוצאות מעניינות על הפערים בין הרכיבים השונים.
תקציר מקורי באנגליתarXiv:2610.03025v1 Announce Type: cross Abstract: What makes a short story gripping; a news article newsworthy; or a math proof elegant? These constructs resist articulation or verification; their meaning is at least partially tacit. However, modern AI models are improved primarily via articulated constitutions, rubrics and verifiers (i.e. in RLAIF and RLVR); tacit components of preferences are typically understudied. We introduce a large, labeled preference dataset CreativePreferences, containing 2.8M texts labeled by 317M human preference judgments across 7 creative domains, with 42 benchmark tasks. We model these labels with executable programs, rubric banks and densely trained models (V, A and VAT, respectively). We observe robust articulability gaps, VAT-VA; and verifiability gaps, VA
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית