כתבה
arXiv cs.AI ·
FineART: מאגר נתונים לניתוח רובוטי
FineART: Fine-Grained Annotated Robotic Trajectory Dataset and Vision-Language-Action Model for Bimanual Manipulation
FineART הוא מאגר נתונים לניתוח רובוטי ברמה עדינה, הכולל 40,543 פרקים ו-533,913 משימות. המאגר מאפשר אימון מודלים לביצוע משימות מורכבות.
תקציר מקורי באנגליתarXiv:2609.36416v2 Announce Type: replace-cross Abstract: Robots operating in real-world environments must often execute complex, multi-step bimanual tasks over long horizons rather than single, isolated actions. Current manipulation datasets struggle to support this capability: although single-arm datasets reach hundreds of thousands of trajectories, they typically provide only one high-level instruction per episode, while existing bimanual datasets with subtask labels annotate only part of their recorded hours. We present FineART, a densely annotated bimanual manipulation dataset comprising 40,543 episodes (1,718 hours) and 533,913 subtasks across 151 tasks. We also introduce FineART-VLA, a vision-language-action policy that predicts its own next subtask to guide its actions. Mid-trainin
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית