יום שישי, 9 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

SP-DocReader: קריאת מסמכים מדויקת יותר

SP-DocReader: Difference-Aware Self-Play for Precise Document OCR
SP-DocReader הוא כלי לקריאת מסמכים מדויקת יותר. הוא משתמש בטכניקת Self-Play כדי לשפר את דיוק הקריאה. התוצאות מראות שיפור של 54% בדיוק הקריאה.
תקציר מקורי באנגליתarXiv:2610.11148v1 Announce Type: cross Abstract: Accurate page transcription remains difficult for vision language models under limited input and training budgets. We present SP-DocReader, a self-play framework for optical character recognition (OCR) that targets residual errors after supervised fine-tuning. Reading Discrepancy Masking aligns reference and generated model tokens through a longest common subsequence, then scores unmatched positions with their full conditioning prefixes. Focused Fidelity Loss adds direct negative log-likelihood supervision at unmatched ground-truth positions. Only the OCR module is trained, while the backbone remains frozen. We derive the combined gradient to distinguish relative score optimization from direct supervision. Compared with SFT-2, SP-DR-3 reduc
קרא במקור המקורי