כתבה
arXiv cs.CL ·
CoDR: טכניקה ללא הכשרה להפחתת דחיפה בביטויי חופשי
CoDR: Training-Free Confidence-Drift Remasking for Diffusion Language Models
CoDR היא טכניקה שאינה דורשת הכשרה להפחתת דחיפה בביטויי חופשי. היא משמשת לשיפור דיוק דגימות של דפוסי חופשי. הטכניקה נבחנה במספר דגימות והוכיחה עצמה כיעילה.
תקציר מקורי באנגליתarXiv:2610.08833v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) decode by repeatedly committing tokens to masked positions, but these commitments are usually irreversible. A token chosen under sparse, partial context is kept fixed, even when later context no longer supports it. Existing samplers mainly decide when to commit a token, but rarely check whether an already committed token should still be kept, allowing early mistakes to propagate. We trace this issue to confidence drift, where the model's confidence in a committed token drops from its sparse commit-time context to the denser context available later. Based on this signal, we propose CoDR (Confidence Drift Remasking), a training-free and sampler-agnostic refinement pass. CoDR estimates drift for all commi
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית