arXiv · 2609.39651
Neural Audio Codec for Robust Audio Deepfake Detection
Abstract
Audio deepfake detectors are typically evaluated on uncompressed audio, although real-world audio often undergoes low-bitrate coding. In this work, we investigate how audio coding affects deepfake detection across codecs, bitrates, and detectors, finding higher errors at lower rates. A mixed-pair protocol isolates codec-induced changes in bona fide and spoof audio, revealing asymmetric, codec-dependent failures: low-rate DAC and EnCodec mainly degrade bona fide detection, whereas X-Codec shows a stronger spoof-side limitation. Motivated by these, we propose a forensic-preserving neural audio codec (FP-NAC), which fine-tunes a pretrained codec using a detector-guided objective while preserving its native hard quantization path and bitrate. On ASVspoof 2019 LA, FP-NAC reduces EER by up to 49.8~pp compared with the original DAC at 0.5~kbps while maintaining comparable reconstruction quality. Although supervised by only one detector, FP-NAC improves performance across multiple detectors, highlighting forensic transparency as a codec design objective alongside perceptual quality. Our codes are available at https://github.com/kjungwoo03/FP-NAC.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Jungwoo Kim, Joonyong Park, Junyoung Koh, Jong-Seok Lee. 2026-09-30. Neural Audio Codec for Robust Audio Deepfake Detection. https://arxiv.org/abs/2609.39651
Cite the original work for its findings. Save a collection to share your selection of sources.