כתבה
arXiv cs.CL ·
תיקון ניבוי, צעדים לא נכונים? גרף ידע קונסנסוס לשינוי סדר ראשון של חשיבה
Correct Prediction, Wrong Steps? Consensus Reasoning Knowledge Graph for Robust Chain-of-Thought Synthesis
תיקון ניבוי, צעדים לא נכונים? חידוש חדש משפר את תהליך החשיבה של ראשון של סדר
תקציר מקורי באנגליתarXiv:2604.14121v5 Announce Type: replace Abstract: Large language models (LLMs) have become increasingly used for various tasks, often coupled with Chain-of-Thought (CoT) prompting to boost accuracy. Recent work has shown that high label-prediction accuracy does not guarantee correct intermediate reasoning, and the causes of *reasoning flaws* vary from sample to sample, yet existing remedies either focus on a single domain or assume that one flaw type applies uniformly across samples. A simple mitigation method is to provide the model with the correct answer, but we show that this yields no consistent improvement in reasoning quality. This indicates that the problem cannot be fixed by LLMs' awareness of answers, and must instead be addressed through the *structure* of reasoning. Motivated
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית