כתבה
arXiv cs.CL ·
Beyond Captions: Context-Grounded Reconstruction for Biomedical Multimodal Continued Pretraining
תקציר מקורי באנגליתarXiv:2606.01049v3 Announce Type: replace Abstract: Biomedical figures are explained not by captions alone but by body-text passages that discuss them. Yet current multimodal corpora typically reduce figures to isolated image-caption pairs, discarding this crucial context. Existing pipelines either omit this context or append it without enforcing the figure references that support each attachment, which can create unsupported image-text attachments and incoherent discourse. We introduce context-grounded reconstruction, a source-grounded framework that converts PubMed Central Open Access (PMC-OA) records into referentially coherent interleaved sequences. It recovers captions and source text, attaches context only through article-native figure references, repairs non-contiguous context, and
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית