arXiv · 2609.03622
Test-time adaptation for speech enhancement with an autoregressive speech prior
Abstract
Test-time adaptation (TTA) offers a promising direction for improving speech enhancement models under mismatched acoustic conditions, without requiring access to labeled target data. In this work, we propose a single-utterance TTA method that regularizes a pretrained speech enhancement model using an autoregressive prior trained on clean speech latent representations extracted from a neural audio codec. Adaptation is performed by minimizing the Kullback-Leibler divergence between the enhanced speech distribution and the clean speech prior. Experiments across multiple noisy speech datasets show consistent improvements in speech quality, particularly under training-testing noise mismatch conditions. Code and audio examples are available online.
Explore related subjects
Keep this discovery
Sofiene Kammoun, Simon Leglaive, Xavier Alameda-Pineda, Timo Gerkmann. 2026-09-03. Test-time adaptation for speech enhancement with an autoregressive speech prior. https://arxiv.org/abs/2609.03622
Cite the original work for its findings. Save a collection to share your selection of sources.
Discover connections
Connections use source metadata and explicit phrase matches, not verified experimental comparisons.