כתבה
arXiv cs.AI ·
מדיניות דיפוזיה ויזואלית-טקטילית שווה-ערכית
Equivariant Visual-Tactile Diffusion Policy for Contact-Rich Manipulation
VISTA היא מדיניות דיפוזיה ויזואלית-טקטילית שווה-ערכית ללמידת חיקוי יעילה. היא משלבת תצפיות ויזואליות וטקטיליות ומנבאת פעולות עקביות. ניסויים מראים ש-VISTA משפרת את יעילות הדגימה.
תקציר מקורי באנגליתarXiv:2610.03333v1 Announce Type: cross Abstract: Imitation learning for contact-rich manipulation requires high-quality expert data that is expensive to obtain. This makes learning a sample-efficient policy a key issue. To address this, we propose VISTA, a workspace-level equivariant visuotactile diffusion policy for data-efficient contact-rich imitation learning. VISTA projects visual and tactile observations into spherical tokens, injects tactile contact cues into visual spherical directions through permutation-equivariant spherical fusion, and rotates the fused harmonic representation using the end-effector orientation. The resulting representation conditions an equivariant diffusion policy to predict spatially consistent actions. Extensive experiments in both simulation and real-world
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית