יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

מה נאמר, ולא מה 'חשב'

What Was Said, Not What Was 'Thought': Type-6 Logic for CoT Verification
המחברים הציגו לוגיקה חדשה לאימות תהליכי החשיבה של LLMs. הלוגיקה, שנקראת Type-6, יכולה לזהות תסכולים וטעויות של LLMs, ולאפשר תיאור ויזואלי של התהליך החשיבתי של המודל. הלוגיקה נבדקה על ידי המחברים על ספריית תרגילים של CoTs, והתוצאות היו מוצלחות.
תקציר מקורי באנגליתarXiv:2609.38420v1 Announce Type: cross Abstract: We introduce Type-6 logic, a variant of dynamic epistemic logic augmented with two operators (uncertainty and recurrence), designed to model the inferential dynamics of contemporary large language model (LLM) chain-of-thought (CoT) reasoning. Type-6 accounts for common LLM reasoning pathologies such as unlicensed revision, enthymemes, loopbacks, and unverifiable/incorrect claims. We propose a verifier based on Type-6 logic that builds a graph out the trace, and checks it against Type-6's axioms and inference rules. We evaluate our framework on LLM-generated CoTs four splits spanning formal and informal reasoning. Our verifier detects structurally unsound reasoning steps that surface-level heuristics miss, and allows for easy visualisation o
קרא במקור המקורי