arXiv · 2605.02928
Keyword spotting using convolutional neural network for speech recognition in Hindi
Abstract
In this study, we investigate the application of keyword spotting (KWS) in the domain of Hindi speech recognition, utilizing a dataset comprising 40,000 audio samples. With a sampling rate of 44 kHz and an average duration of 1.9 seconds per sample, we focus on developing an efficient on-device KWS system tailored for user-specific queries. Leveraging Convolutional Neural Networks (CNNs) for classification, we employ feature engineering techniques to convert raw audio recordings into Mel Frequency Cepstral Coefficients (MFCCs) as an input for our network. Our experiments encompass various CNN architectures, exploring their efficacy in identifying predefined keywords within the continuous speech stream. Our CNN-based approach achieves a commendable accuracy rate of 91.79% through rigorous evaluation, demonstrating promising performance while ensuring computational efficiency and user-specific customization in Hindi speech recognition.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Saru Bharti, Pushparaj Mani Pathak. 2026-04-26. Keyword spotting using convolutional neural network for speech recognition in Hindi. https://doi.org/10.1109/icccnt61001.2024.10724682
Cite the original work for its findings. Save a collection to share your selection of sources.