כתבה
arXiv cs.AI ·
מדידת אחריות AI דרך ביקורת על טיעונים: האם תקינות התקינות של ה-LLM עומדת במבחן?
Measuring AI Accountability Through Argumentation Analysis: Can Model Reasoning Withstand Scrutiny?
במאמר זה נחקר כיצד לבדוק את האחריות של LLMs דרך ביקורת על טיעונים. המחברים פיתחו פרוטוקול דיאלקטי שמאפשר לבדוק את התקינות של ה-LLM.
תקציר מקורי באנגליתarXiv:2609.05088v1 Announce Type: new Abstract: AI oversight methods rely on ground truth for validation, but what constitutes appropriate AI behavior is contested. This leaves evaluation of moral reasoning in LLMs and debate-based oversight implicitly avoiding realistic ambiguity. We investigate an alternative standard designed to function despite such ambiguity: structural quality of the defence a model can mount for its verdicts in response to critical questions, measured through a four-phase dialectical protocol grounded in Walton's theory of argumentation schemes and Govier's criteria for argument cogency. The protocol is adaptive to different frames of reasoning, extends beyond multiple-choice framing, and treats both the reasoning that precedes a verdict and its post-hoc justificati
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית