arXiv · 1910.05603
VAIS ASR: Building a conversational speech recognition system using language model combination
Abstract
Automatic Speech Recognition (ASR) systems have been evolving quickly and reaching human parity in certain cases. The systems usually perform pretty well on reading style and clean speech, however, most of the available systems suffer from situation where the speaking style is conversation and in noisy environments. It is not straight-forward to tackle such problems due to difficulties in data collection for both speech and text. In this paper, we attempt to mitigate the problems using language models combination techniques that allows us to utilize both large amount of writing style text and small number of conversation text data. Evaluation on the VLSP 2019 ASR challenges showed that our system achieved 4.85% WER on the VLSP 2018 and 15.09% WER on the VLSP 2019 data sets.
Explore related subjects
Keep this discovery
Quang Minh Nguyen, Thai Binh Nguyen, Ngoc Phuong Pham, The Loc Nguyen. 2019-10-12. VAIS ASR: Building a conversational speech recognition system using language model combination. https://arxiv.org/abs/1910.05603
Cite the original work for its findings. Save a collection to share your selection of sources.