יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.CL ·

Ex-Omni: ייצור אנימציה תלת-ממדית

Ex-Omni: Enabling 3D Facial Animation Generation for Omni-modal Large Language Models
Ex-Omni הוא כלי לייצור אנימציה תלת-ממדית של פנים, המשלב מודלים של שפה ותנועה. הוא מאפשר ייצור תנועות פנים מתואמות עם דיבור.
תקציר מקורי באנגליתarXiv:2602.07106v3 Announce Type: replace-cross Abstract: Omni-modal large language models (OLLMs) aim to unify multimodal understanding and generation, yet extending them to jointly produce speech and 3D facial animation remains largely underexplored. A key challenge is the mismatch between the discrete semantic reasoning of LLMs and the dense temporal dynamics required for 3D facial motion. We propose Expressive Omni (Ex-Omni), a framework that augments OLLMs with speech-accompanied 3D facial animation. Ex-Omni decouples semantic reasoning from temporal generation through a speech-unit generator with blendshape co-supervision and a non-autoregressive blendshape decoder, where speech units provide temporal scaffolding and hidden speech representations carry facially relevant cues. We furt
קרא במקור המקורי