arXiv · 2305.16043
Ordered and Binary Speaker Embedding
Abstract
Modern speaker recognition systems represent utterances by embedding vectors. Conventional embedding vectors are dense and non-structural. In this paper, we propose an ordered binary embedding approach that sorts the dimensions of the embedding vector via a nested dropout and converts the sorted vectors to binary codes via Bernoulli sampling. The resultant ordered binary codes offer some important merits such as hierarchical clustering, reduced memory usage, and fast retrieval. These merits were empirically verified by comprehensive experiments on a speaker identification task with the VoxCeleb and CN-Celeb datasets.
Explore related subjects
Keep this discovery
Jiaying Wang, Xianglong Wang, Namin Wang, Lantian Li, Dong Wang. 2023-05-25. Ordered and Binary Speaker Embedding. https://arxiv.org/abs/2305.16043
Cite the original work for its findings. Save a collection to share your selection of sources.