יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

FaithfulBench: האם עוזרי AI תומכים או פוגעים באמונת המשתמש?

FaithfulBench: Does AI Counsel Uphold or Undermine the User's Professed Faith?
עוזרי AI תומכים או פוגעים באמונת המשתמש? ניתוח של FaithfulBench, הבנצ'מרק הראשון לבדיקת עוזרי AI בקשר לאמונות של המשתמש. נבדקו חמשה עוזרי AI חדשים, והתגלה שהם פוגעים באמונות של חלק מהמשתמשים.
תקציר מקורי באנגליתarXiv:2609.13634v1 Announce Type: cross Abstract: Do AI assistants help believers reason about moral dilemmas consistently with their faith? We present FaithfulBench, the first benchmark to score AI counsel across traditions by how well it adheres to the user's professed faith. Scenarios are drawn from each tradition's most respected texts, with the faithful answer known and applied by the judges as the standard. We test five frontier models under three conditions: the AI does not know the user's tradition; it receives a one-line prompt identifying the user as a practicing adherent; or it receives a companion-counselor guide rooted in the tradition's sources. Two judges score the initial response and whether the model caves or holds when pressured toward the answer the user wants. When the
קרא במקור המקורי