יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

איחוד של תפעולי וידאו דרך אנלוגיה ספאציו-טמפורלית

Unifying Video Tasks via Spatiotemporal Analogy
אנלוגיה ספאציו-טמפורלית לאיחוד של תפעולי וידאו. ViGeo, פלטפורמה שמשלבת למידה במצבי-מצורף וגרפים ספאציו-טמפורליים, מאפשרת תפעולי וידאו חדשים תוך שימוש בקלט וידאו קיים.
תקציר מקורי באנגליתarXiv:2609.33935v2 Announce Type: replace-cross Abstract: Adapting video models to new tasks typically requires dedicated data curation and fine-tuning. While visual analogy provides a training-free alternative by specifying tasks in-context, it remains restricted to the image domain. To explore whether analogy-based methods can unify diverse video tasks and generalize to out-of-distribution scenarios, we introduce ViGeo, a framework that extends visual in-context learning to the video domain via spatiotemporal canvas completion. Evaluated on a diverse task taxonomy with a strict train-test split, ViGeo generalizes to unseen video manipulations and zero-shot modalities (e.g., event cameras). Finally, we identify task internalization, where a query format associated with a pretrained task o
קרא במקור המקורי