כתבה
arXiv cs.AI ·
LayerRoute: Action-Conditioned Mixture-of-Layers Routing for Vision-Language-Action Policies
תקציר מקורי באנגליתarXiv:2609.06079v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies leverage pretrained vision-language models (VLMs) to guide action generation for robot control. VLMs provide hierarchical visual-semantic representations that evolve across layers, from local visual geometry to abstract, language-aligned semantics; different manipulation tasks may therefore require different mixtures of layer representations. Meanwhile, the action module maintains intermediate representations that evolve throughout action computation and may provide useful information for subsequent decisions. However, existing VLA interfaces offer limited flexibility in representation access: VLM information is exposed through fixed layer assignments for each action layer, while intermediate action state
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית