arXiv · 2312.15190
SAIC: Integration of Speech Anonymization and Identity Classification
Abstract
Speech anonymization and de-identification have garnered significant attention recently, especially in the healthcare area including telehealth consultations, patient voiceprint matching, and patient real-time monitoring. Speaker identity classification tasks, which involve recognizing specific speakers from audio to learn identity features, are crucial for de-identification. Since rare studies have effectively combined speech anonymization with identity classification, we propose SAIC - an innovative pipeline for integrating Speech Anonymization and Identity Classification. SAIC demonstrates remarkable performance and reaches state-of-the-art in the speaker identity classification task on the Voxceleb1 dataset, with a top-1 accuracy of 96.1%. Although SAIC is not trained or evaluated specifically on clinical data, the result strongly proves the model's effectiveness and the possibility to generalize into the healthcare area, providing insightful guidance for future work.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Ming Cheng, Xingjian Diao, Shitong Cheng, Wenjun Liu. 2023-12-23. SAIC: Integration of Speech Anonymization and Identity Classification. https://arxiv.org/abs/2312.15190
Cite the original work for its findings. Save a collection to share your selection of sources.