כתבה
arXiv cs.AI ·
VOMMI: איסוף וניצול הדגמאות ניידות למניפולציה
VOMMI: Collecting and Leveraging Portable Demonstrations for Mobile Manipulation
VOMMI הוא כלי לאיסוף וניצול הדגמאות ניידות למניפולציה. הוא משתמש ב-RGB כדי ללמוד מהדגמאות ולשפר את היכולת של רובוטים לבצע משימות. VOMMI משלב טראוריה ויזואלית עם פעולות ותנועות, ומאפשר לרובוטים ללמוד מניסיון בצורה יותר יעילה.
תקציר מקורי באנגליתarXiv:2610.08220v1 Announce Type: cross Abstract: Portable mobile-manipulation demonstrations can help alleviate data scarcity for embodied intelligence, but obtaining reliable, low-cost, and robot-free motion supervision from RGB observations remains challenging. Existing approaches often rely on teleoperation or specialized devices equipped with additional sensing hardware, while directly using estimated visual odometry (VO) trajectories can introduce inconsistencies due to accumulated drift and imperfect motion supervision. We present the Visual-Odometry-Conditioned Mobile Manipulation Interface (VOMMI), a portable demonstration collection and learning framework that connects portable RGB demonstrations to vision-language-action (VLA) post-training through offline trajectory reconstruct
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית