יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

ראייה אינה על-גובה: יישור חזק של קבצים באחד-עברית ללא-הפסדי דקודינג מוציא-טריפ במודלי ראייה-שפה

Vision Is Not Overhead: One-Pass Block Drafting for Lossless Speculative Decoding in Vision-Language Models
דקודינג מוציא-טריפ מקדם ייצור ללא שינוי בתוצאות, אך במודלי ראייה-שפה, ציקלון עצמאי מחזיק אותו בחזקה.
תקציר מקורי באנגליתarXiv:2609.00355v2 Announce Type: replace Abstract: Speculative decoding accelerates generation without changing its output, but on vision-language models (VLMs) a self-reinforcing cycle holds it back. Because an autoregressive drafter pays a sequential pass for each drafted token, it must stay small and can ill afford to attend to the image at each pass. Prior work therefore compresses or hides the image, leaving the drafter weakest on the text the image determines. We present GLANCE, a one-pass block drafter that breaks this cycle on an unmodified VLM target. Its block-diffusion head drafts a whole block in one forward pass over the target's already fused vision-language states, reading the multimodal context once, however deep the draft. The target verifies a wide candidate tree in one
קרא במקור המקורי