כתבה
arXiv cs.LG ·
TokenMapper: צעד כלפי תקשורת רציפה בין מודלי דיבור
TokenMapper: A Step Toward Interoperable Speech Token Translation
TokenMapper מציע פתרון לבעיה של תקשורת רציפה בין מודלי דיבור שונים. הפתרון, TokenMapper, מאפשר תרגום ישיר בין תווי דיבור שונים, כולל תווי דיבור עם קודבוק יחיד ותווי דיבור עם קודבוק רב-ערכי. TokenMapper נבחן על ידי ניסויים שהראו תרגום יעיל ובעל תוצאות טובות.
תקציר מקורי באנגליתarXiv:2609.12563v1 Announce Type: new Abstract: Neural audio codecs discretize speech into token sequences, but the resulting token spaces differ in vocabulary and codebook structure, preventing direct communication across models. This limitation affects applications such as conversational voice agents and speech to speech translation systems where multiple speech models must interact. As a result, transferring information between speech systems typically requires decoding to waveform audio and re-encoding with a second tokenizer, increasing latency and introducing potential information loss. To address these limitations, we present TokenMapper, a direction aware framework for direct token to token translation between heterogeneous speech tokenizers in the discrete domain. TokenMapper supp
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית