כתבה
arXiv cs.LG ·
בדיקות סיכוך-מסומן לסיבות תלויות-תשובה במודלי LLM
Context-Masked Truncated Reasoning Audits for Answer-Key Dependence in LLM Tutors
בדיקות סיכוך-מסומן לסיבות תלויות-תשובה במודלי LLM. נמצא כי חשיפה לתשובה ישירה גורמת לשיפור בדיוק של המודל. נכללו גם ניסויים שבהם נסתרה התשובה והמודל נדרש לסקור את התשובה.
תקציר מקורי באנגליתarXiv:2607.04572v3 Announce Type: replace-cross Abstract: Large language model (LLM) tutors may have access to teacher notes, answer keys, rubrics, or retrieved solutions while producing student-facing explanations. We study whether truncated reasoning probes can distinguish direct access to such private context from answer information carried by the written explanation. Using Truncated Reasoning AUC Evaluation (TRACE), we evaluate 1000 GSM8K problems under question-only, correct answer-key, and wrong answer-key contexts. When forced-answer probes retain the private key, answer-key TRACE AUC rises from 0.375 to 0.900, and the gold answer is recoverable with no explanation at all in 998 of 1000 cases. We then introduce a context-masked replay: answer-key-generated prefixes are probed under
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית