כתבה
arXiv cs.LG ·
ניהול ושירות ראייה-שפה-פעולה יעילים למפעלי רובוטים
Efficient Vision-Language-Action Management and Serving for Robot Factories
מערכת ניהול ושירות ראייה-שפה-פעולה יעילים למפעלי רובוטים. המערכת, שנקראת Robion, מסוגלת לשרת עד 64 רובוטים ב-98% סיכוי להשגת סטנדרטי תפעול (SLO).
תקציר מקורי באנגליתarXiv:2609.12075v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models show high robotic manipulation capabilities via a two-stage design: a Vision-Language Model (VLM) stage followed by an Action Diffusion Transformer (ADiT) stage. Since robots must meet strict Service-Level Objectives (SLOs) for safety, VLA inference is inherently latency-critical. Meeting these SLOs requires high-end GPUs, yet weight, cost, and power constraints preclude integrating such GPUs on-robot. Prior works offload VLA inference to edge servers that serve many robots on VLA models. However, current VLA systems lack support for multi-request, multi-model execution on a multi-GPU server under SLOs, while existing serving systems for multi-stage models are optimized for throughput and stage disaggrega
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית