כתבה
arXiv cs.CL ·
EmoRES-TTS: Residual-Enhanced Vector Steering for Emotional Speech Generation
תקציר מקורי באנגליתarXiv:2609.38157v1 Announce Type: cross Abstract: Emotion-conditioned text-to-speech (TTS) models may fail to express the requested emotion reliably, and improving controllability by additional training is costly in both computation and emotion-labeled speech training data. We therefore study vector steering, a training-free approach that modifies the internal representations of a frozen model. CoCoEmo, a conventional vector steering method for emotion TTS, treats each emotion vector as an indivisible direction controlled by a single global strength, limiting adherence to the requested emotion. In this work, we first discover that an emotion vector can be decomposed into a shared component that moves speech away from neutral expression and a residual component that directs generation towar
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית