כתבה
arXiv cs.AI ·
Text-to-3D Policy: Fine-Grained Language-Behavior Alignment for Unseen Specification Generalization
תקציר מקורי באנגליתarXiv:2609.39599v1 Announce Type: cross Abstract: 3D visuomotor policies provide a strong foundation for spatially precise manipulation, yet current text-to-3D policies struggle to follow unseen fine-grained behavioral specifications beyond those covered by demonstrations. We study this challenge as unseen specification generalization, where language specifies behaviorally significant variations, such as target position, displacement, or articulated state, that are absent from policy training. We find that pretrained language representations and conventional global behavior-language alignment capture coarse task semantics but often blur nearby specifications that require distinct behaviors. We introduce T3DP, a Text-to-3D Policy framework for fine-grained language-behavior alignment. Rathe
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית