כתבה
arXiv cs.CL ·
מומחיות מוסרית לפני תוכן מוסרי: למה סוגיות LLM חסרות את התנאים הנדרשים להתאמה עדכנית
Moral Competence Before Moral Content: Why LLM Agents Lack the Prerequisites for Coherent Alignment
סוגיות LLM חסרות את התנאים הנדרשים להתאמה עדכנית, כתוצאה מאי-יציבות פסקנים וחוסר מומחיות מוסרית. המחברים הציגו ארבעה תנאים מבני למדיניות עדכנית, והדגימו את השיטה בשלושה סימולציות של סוגיות LLM. התוצאות הראו שאף סוגיית LLM לא הציגה מדיניות עדכנית על פני שלוש הסימולציות.
תקציר מקורי באנגליתarXiv:2609.05036v1 Announce Type: cross Abstract: AI alignment requires AI systems to adhere to human norms, values, or intentions. Under value pluralism there is no correct target, but a shared prerequisite is that the system's behavior expresses a coherent policy: a mapping from situations to verdicts that is invariant while a situation's morally relevant features are preserved, and sensitive when they change. We introduce four structural conditions for such coherent policies: verdict stability, monotonicity, decisiveness, and Pareto viability. Together they measure a form of moral competence that is evaluable from behavior alone, without reference to a moral standard or expert baseline, forming a structural floor for alignment rather than a normative target. We demonstrate the methodolo
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית