כתבה
arXiv cs.AI ·
VINCIE-NExT: פתיחת עריכת וידאו מתוך תמונות באמצעות דגמי תצוגה
VINCIE-NExT: Unlocking Video Editing from Images via In-Context Modeling
VINCIE-NExT היא פלטפורמה שמאפשרת עריכת וידאו מתוך תמונות באמצעות דגמי תצוגה. הפלטפורמה משתמשת בשיטת התפשטות כדי להעביר יכולות עריכה מתמונות לווידאו. VINCIE-NExT חולקה לשלבי עריכה שונים, כאשר כל שלב מורכב משלבי עריכה קטנים יותר. הפלטפורמה גם מאפשרת עריכה של וידאו באמצעות תמונות שנוצרו על ידי המודל או על ידי המשתמש.
תקציר מקורי באנגליתarXiv:2610.12104v1 Announce Type: cross Abstract: Building a capable video editor remains significantly harder than a video generator: editing requires (source, instruction, edited) triplets that are prohibitively expensive to annotate and difficult to synthesize at scale, whereas image editing has already reached maturity with millions of such pairs readily available. In this work, we introduce VINCIE-NExT, a unified framework that transfers editing capability from images to videos through in-context visual demonstrations, alleviating the need for large-scale paired video editing data. VINCIE-NExT decomposes video editing into a structured chain of composable sub-tasks (Video -> Image -> Image -> Video), routing editing intent through the image domain and enabling scalable joint training
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית