כתבה
arXiv cs.CL ·
פירוק מידע לשוני ופרה-לשוני עם Routed Sparse Autoencoders
Disentangling Linguistic and Paralinguistic Information with Routed Sparse Autoencoders
חוקרים מציגים שיטה לפירוק מידע לשוני ופרה-לשוני באמצעות Routed Sparse Autoencoders. השיטה מאפשרת הפרדה ברורה בין מידע לשוני למידע פרה-לשוני, כולל זהות הדובר, רגשות ופרוזודיה.
תקציר מקורי באנגליתarXiv:2610.10865v1 Announce Type: new Abstract: Self-supervised speech encoders contain linguistic and paralinguistic information in a shared, entangled representation space. We combine a TopK sparse autoencoder with route-specific supervision and cross-factor adversaries. Across frozen SPEAR and WavLM encoders, independent probes show factor-specific retention and suppression: linguistic information remains stronger in the linguistic route, while paralinguistic factors, including speaker identity, emotion, and prosody, are retained in the paralinguistic route and substantially reduced in the linguistic route. The route organisation learned on LibriSpeech persists on MSP-Podcast without representation-side retraining. Feature-space route interventions further transfer the swapped factor wh
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית