arXiv · 2412.18955
Leave-One-EquiVariant: Alleviating invariance-related information loss in contrastive music representations
Abstract
Contrastive learning has proven effective in self-supervised musical representation learning, particularly for Music Information Retrieval (MIR) tasks. However, reliance on augmentation chains for contrastive view generation and the resulting learnt invariances pose challenges when different downstream tasks require sensitivity to certain musical attributes. To address this, we propose the Leave One EquiVariant (LOEV) framework, which introduces a flexible, task-adaptive approach compared to previous work by selectively preserving information about specific augmentations, allowing the model to maintain task-relevant equivariances. We demonstrate that LOEV alleviates information loss related to learned invariances, improving performance on augmentation related tasks and retrieval without sacrificing general representation quality. Furthermore, we introduce a variant of LOEV, LOEV++, which builds a disentangled latent space by design in a self-supervised manner, and enables targeted retrieval based on augmentation related attributes.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Julien Guinot, Elio Quinton, György Fazekas. 2024-12-25. Leave-One-EquiVariant: Alleviating invariance-related information loss in contrastive music representations. https://arxiv.org/abs/2412.18955
Cite the original work for its findings. Save a collection to share your selection of sources.