כתבה
arXiv cs.CL ·
Content Anonymization for Privacy in Long-form Audio
תקציר מקורי באנגליתarXiv:2510.12780v3 Announce Type: replace-cross Abstract: Voice anonymization techniques have been found to successfully obscure a speaker's acoustic identity in short, isolated utterances in benchmarks such as the VoicePrivacy Challenge. In practice, however, utterances seldom occur in isolation: long-form audio is commonplace in domains such as interviews, phone calls, and meetings. In these cases, many utterances from the same speaker are available, which pose a significantly greater privacy risk: given multiple utterances from the same speaker, an attacker could exploit an individual's vocabulary, syntax, and turns of phrase to re-identify them, even when their voice is completely disguised. To address this risk, we propose a new approach that performs a contextual rewriting of the tra
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית