arXiv · 2508.18907
SegReConcat: A Data Augmentation Method for Voice Anonymization Attack
Abstract
Anonymization of voice seeks to conceal the identity of the speaker while maintaining the utility of speech data. However, residual speaker cues often persist, which pose privacy risks. We propose SegReConcat, a data augmentation method for attacker-side enhancement of automatic speaker verification systems. SegReConcat segments anonymized speech at the word level, rearranges segments using random or similarity-based strategies to disrupt long-term contextual cues, and concatenates them with the original utterance, allowing an attacker to learn source speaker traits from multiple perspectives. The proposed method has been evaluated in the VoicePrivacy Attacker Challenge 2024 framework across seven anonymization systems, SegReConcat improves de-anonymization on five out of seven systems.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Ridwan Arefeen, Xiaoxiao Miao, Rong Tong, Aik Beng Ng, Simon See. 2025-08-26. SegReConcat: A Data Augmentation Method for Voice Anonymization Attack. https://arxiv.org/abs/2508.18907
Cite the original work for its findings. Save a collection to share your selection of sources.