כתבה
arXiv cs.CL ·
ווקטורים מחליפים הדגמות
Two Vectors Replace In-Context Demos: Structured Task Adaptation via Embeddings
STAVE מחליף הדגמות בווקטורים ספציפיים למשימה. השיטה משתמשת בווקטורים לעדכון קבוצות טוקנים. ניסויים על מודלים רב-מודאליים ומודלי שפה גדולים הראו תוצאות טובות.
תקציר מקורי באנגליתarXiv:2610.07572v1 Announce Type: new Abstract: In-context learning (ICL) adapts frozen large multimodal models (LMMs) to new tasks from a few demonstrations (demos), but re-encodes them at every query, where each demo image adds up to hundreds of visual tokens. Demo-free methods remove this cost with a compact task state. However, they add it at locations searched per task or at every decoder layer, where task parameters grow with depth. Moreover, inserted tokens or keys cannot change how the original prompt divides its attention within a layer. To address these issues, we propose Structured Task Adaptation via Embeddings (STAVE), which replaces demos with two task-specific vectors added to existing input embeddings. Specifically, a readout vector updates the answer-producing tokens and a
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית