כתבה
arXiv cs.AI ·
העברת תרגום בין סימולציה לעולם האמיתי לשיטת ניווט תצלום-לשפה
Sim-to-Real Transfer of Vision-Language Navigation in Continuous Environments Using an Ackermann-Steered Mobile Robot
במאמר זה, נציגים תרגום סימולציה-עולם-אמיתי לשיטת ניווט תצלום-לשפה, שמשתמשת ברובוט ניווטי עם נעה-אקרמן. השיטה משתמשת ב-Cross-Modal Attention (CMA) ומצלמה ו-LiDAR לניווט באזורים רציפים.
תקציר מקורי באנגליתarXiv:2610.07192v1 Announce Type: new Abstract: Vision-Language Navigation (VLN) enables robots to navigate through environments using natural language instructions, making human-robot interaction intuitive. Traditional VLN models often rely on navigation graphs, 360-degree views, and perfect localization which pose significant challenges when adapting these models to real-world settings. This work addresses these limitations by performing a simulation-to-real domain shift of a VLN approach that operates in continuous environments without requiring navigation graphs or panoramic views. The proposed system integrates vision-language models that align visual inputs and linguistic instructions within a shared embedding space, facilitating natural language-driven navigation. We employ a Cross-
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית