כתבה
arXiv cs.AI ·
LightNav-0: גילוי תבונה חללית לניווט גוף-כללי
LightNav-0: Eliciting VLM Spatial Intelligence for Generalist Embodied Navigation
LightNav-0 מגלה תבונה חללית לניווט גוף-כללי, כולל גילוי תבונה חללית של VLM, ואלימות חללית של VLM. המודל מייצג תפקידים שונים באמצעות פונטינג דו-קנלי, ומפנה זיהוי וסימון חללית. LightNav-ER, הביצועים המגובשים של LightNav-0, הגיעו לשיא קבוצת הביצועים המלאה ב-8 ביצועים-מחשבים, ו-LightNav-0 הגיע לשיא ביצועים-מחשבים ב-10 סימולציות-ניווט ציבוריות.
תקציר מקורי באנגליתarXiv:2608.30935v2 Announce Type: replace-cross Abstract: Embodied navigation requires agents to translate heterogeneous goals and visual observations into actions across tasks, environments, and robot embodiments. Modern vision-language models (VLMs) already encode spatial priors for visual grounding, spatial reasoning, and pointing, but these capabilities are rarely elicited directly for robot control. Existing navigation systems instead rely on task- or embodiment-specific components, fragmenting perception, reasoning, and action while offering limited generalization. Here we present LightNav-0, a compact generalist embodied navigation model that elicits the spatial intelligence of a pretrained VLM and aligns it with navigation, without task-specific prediction heads. LightNav-0 represe
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית