יום שישי, 31 ביולי 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

MarineEVT: התקדמות בהבנת אירועים-מרכזיים בווידאו ימי באמצעות תוכן ראיון ויזואלי

MarineEVT: Advancing Event-Centric Marine Video Understanding via Visual Tool Reasoning
נוצרה סטטיסטיקה ימית חדשה, MarineEVT, ומודל EVT-R1, המספקים הבנה יוצאת דופן של וידאו ימי, על ידי התמקדות באירועים-מרכזיים.
תקציר מקורי באנגליתarXiv:2607.24064v1 Announce Type: cross Abstract: Recent Vision-Language Models (VLMs) have achieved remarkable success in visual understanding, driven by the growing availability of high-quality image-text pairs. However, the performance of VLMs often degrades in the video domain due to the essential need for temporal understanding and the scarcity of large-scale annotated video data. In this work, we focus on marine video understanding, which brings further challenges: first, it requires substantial domain expertise; and video VLMs usually struggle with localizing and interpreting critical information from marine videos, as the informative events are typically sparse, unpredictable, and unevenly distributed. To address these challenges, we carefully curate the first event-centric marine
קרא במקור המקורי