arXiv · 2508.08559
Multi-Target Backdoor Attacks Against Speaker Recognition
Abstract
In this work, we propose a multi-target backdoor attack against speaker identification using position-independent clicking sounds as triggers. Unlike previous single-target approaches, our method targets up to 50 speakers simultaneously, achieving success rates of up to 95.04%. To simulate more realistic attack conditions, we vary the signal-to-noise ratio between speech and trigger, demonstrating a trade-off between stealth and effectiveness. We further extend the attack to the speaker verification task by selecting the most similar training speaker - based on cosine similarity - as a proxy target. The attack is most effective when target and enrolled speaker pairs are highly similar, reaching success rates of up to 90% in such cases.
Explore related subjects
Keep this discovery
Alexandrine Fortier, Sonal Joshi, Thomas Thebaud, Jesús Villalba, Najim Dehak, Patrick Cardinal. 2025-08-12. Multi-Target Backdoor Attacks Against Speaker Recognition. https://arxiv.org/abs/2508.08559
Cite the original work for its findings. Save a collection to share your selection of sources.