כתבה
arXiv cs.AI ·
Harnessing Coupled Stream Completion For Human-Object Interaction Modeling
תקציר מקורי באנגליתarXiv:2609.32551v2 Announce Type: replace-cross Abstract: Text-conditioned human-object interaction (HOI) generation requires body motion, object trajectories & rotations, and hand articulation to remain coordinated. These components differ in scale and dynamics, but must agree on contact, relative pose, and timing. A shared representation may limit the distinct structure of each stream, while independent generation prevents each stream from responding to changes in the others. Latent supervision alone also does not directly constrain contact after decoding. We propose TRACE, a continuous latent framework that keeps stream states separate and couples their updates. TRACE encodes body, object, and hand motion into separate latents and predicts each stream velocity from the complete current
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית