כתבה
arXiv cs.AI ·
ניווט הומנואידי בסביבות מורכבות
TANGO: Humanoid Navigation in Cluttered Environments with a Whole-Body Vision-Language-Action Model
TANGO הוא מודל ניווט הומנואידי שמאפשר ניווט בסביבות מורכבות. המודל משתמש בראייה, שפה ופעולה כדי לנווט בחללים תלת-ממדיים. TANGO מאומן באמצעות סימולציה ומציג ביצועים מצוינים בניסויים.
תקציר מקורי באנגליתarXiv:2609.09158v2 Announce Type: replace-cross Abstract: We study the problem of navigating cluttered indoor environments with a humanoid robot. Unlike conventional methods that model navigation as a 2D path planning problem, humanoid traversal in cluttered environments requires continuous geometry-aware whole-body adaptation, including coordinated arm placement, torso adjustment, and gait modulation for collision-free movement through complex 3D spaces. We introduce TANGO, the first whole-body vision-language navigation framework for language-conditioned humanoid traversal in cluttered environments. Given a natural-language instruction and egocentric RGB observations, TANGO directly predicts 29-DoF joint-space actions for downstream whole-body control. We train TANGO entirely in simulati
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית