כתבה
arXiv cs.AI ·
תכונה מוסרית לפני תוכן מוסרי
Moral Competence Before Moral Content: Why LLM Agents Lack the Prerequisites for Coherent Alignment
חוקרים בדקו את היכולת של מודלים LLM להתנהג באופן מוסרי. הם מצאו שאף מודל לא הצליח להראות תכונה מוסרית עקבית. המחקר בוחן ארבע תנאים מבניים לתכונה מוסרית עקבית.
תקציר מקורי באנגליתarXiv:2609.05036v1 Announce Type: new Abstract: AI alignment requires AI systems to adhere to human norms, values, or intentions. Under value pluralism there is no correct target, but a shared prerequisite is that the system's behavior expresses a coherent policy: a mapping from situations to verdicts that is invariant while a situation's morally relevant features are preserved, and sensitive when they change. We introduce four structural conditions for such coherent policies: verdict stability, monotonicity, decisiveness, and Pareto viability. Together they measure a form of moral competence that is evaluable from behavior alone, without reference to a moral standard or expert baseline, forming a structural floor for alignment rather than a normative target. We demonstrate the methodology
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית