arXiv · 2608.24558
Array-Agnostic Ambisonics Encoding via Diffusion Posterior Sampling
Abstract
Spatial audio enhances user immersion by reproducing 3D sound fields, with Ambisonics being a widely adopted representation. While Ambisonics is theoretically independent of the recording setup, practical microphone arrays introduce hardware-dependent encoding artifacts. Moreover, existing data-driven solutions lack flexibility, as they are typically restricted to fixed array geometries. To overcome these limitations, we propose ADEPS, a generative framework that explicitly embeds the physical acquisition model into the inference process. By leveraging this formulation, ADEPS effectively compensates for array-specific distortions while enabling zero-shot encoding across arbitrary array topologies. We train the underlying generative prior in an unsupervised manner solely on target Ambisonic representations. Extensive evaluations across diverse simulated and real microphone arrays demonstrate that ADEPS consistently outperforms both traditional linear and parametric baselines in spatial fidelity and spectral quality.
Explore related subjects
Keep this discovery
Amit Milstein, Nir Shlezinger, Boaz Rafaely. 2026-08-25. Array-Agnostic Ambisonics Encoding via Diffusion Posterior Sampling. https://arxiv.org/abs/2608.24558
Cite the original work for its findings. Save a collection to share your selection of sources.