יום שישי, 9 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

IndustrialVLA-Bench: כלי חדש לבדיקת דגמי מדיניות רובוטים

IndustrialVLA-Bench: A Traceable Multi-Axis Evaluation of Open Robot Policy Models
כלי חדש לבדיקת דגמי מדיניות רובוטים, המספק תוצאות עדכניות ומדויקות לדגמי VLA ו-WAM.
תקציר מקורי באנגליתarXiv:2609.25562v2 Announce Type: replace-cross Abstract: Open robot policies increasingly follow two paradigms: vision-language-action models (VLAs) directly map observations and instructions to actions, whereas world-action models (WAMs) incorporate learned video or world dynamics into policy learning or action generation. Although both target the same manipulation tasks and represent alternative design choices, they are commonly reported under different evaluation protocols, leaving their capability, robustness, language sensitivity, and deployment-cost trade-offs unclear. We present IndustrialVLA-Bench, an evidence-aware evaluation of six released VLA and WAM systems under a unified reporting schema. It separately evaluates clean capability on LIBERO, non-language robustness on LIBERO-
קרא במקור המקורי