arXiv · 2605.17443
Analyzing Error Propagation in Korean Spoken QA with ASR-LLM Cascades
Abstract
We analyze how automatic speech recognition (ASR) errors propagate through ASR--LLM cascades in Korean spoken question answering (SQA), focusing on downstream semantic failures that conventional ASR metrics cannot fully capture. Our analysis shows that the relative downstream degradation caused by ASR errors is consistent across LLMs with different absolute performance, suggesting that cascade degradation largely tracks ASR-stage information loss. We further identify single-character ASR errors as a particularly salient source of information loss in Korean, where even a minimal transcription difference can change the intended question and degrade downstream QA performance. Finally, an auxiliary comparison shows that a large audio language model outperforms an ASR--LLM cascade with an approximately matched language backbone in noisy Korean SQA, indicating the potential of direct audio input to mitigate transcript-induced information loss.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Donghyuk Jung, Youngwon Choi. 2026-05-17. Analyzing Error Propagation in Korean Spoken QA with ASR-LLM Cascades. https://arxiv.org/abs/2605.17443
Cite the original work for its findings. Save a collection to share your selection of sources.