יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

איך Lossless הוא Lossless Speculative Decoding?

How Lossless Is Lossless Speculative Decoding? The Role of Numerical Precision in Orthrus
Orthrus הוא ארכיטקטורה היברידית שמאיצה את היכולת של מודלי שפה אוטורגרסיביים. המחקר בודק את הטענה ש-Orthrus מאפשר פענוח ספקולטיבי ללא הפסדים. התוצאות מראות שהדבר תלוי ברמת הדיוק הנומרית.
תקציר מקורי באנגליתarXiv:2609.15504v1 Announce Type: cross Abstract: Orthrus is a hybrid autoregressive-diffusion architecture that accelerates autoregressive language-model inference by generating multiple tokens in parallel while using a frozen autoregressive backbone. Its central claim is that an intra-model consensus mechanism enables lossless speculative decoding, producing the same output sequence as the autoregressive model. We independently reproduce Orthrus and examine this claim under different numerical precisions. Under BF16 inference, exact trajectory matching occurs in only 45% of cases for the authors' checkpoint and 43% for our independently trained model across 1,190 prompts from 12 domains. The probability of exact matching is also strongly associated with the response-conditional perplexit
קרא במקור המקורי