arXiv · 2304.08431
Prak: An automatic phonetic alignment tool for Czech
Abstract
Labeling speech down to the identity and time boundaries of phones is a labor-intensive part of phonetic research. To simplify this work, we created a free open-source tool generating phone sequences from Czech text and time-aligning them with audio. Low architecture complexity makes the design approachable for students of phonetics. Acoustic model ReLU NN with 56k weights was trained using PyTorch on small CommonVoice data. Alignment and variant selection decoder is implemented in Python with matrix library. A Czech pronunciation generator is composed of simple rule-based blocks capturing the logic of the language where possible, allowing modification of transcription approach details. Compared to tools used until now, data preparation efficiency improved, the tool is usable on Mac, Linux and Windows in Praat GUI or command line, achieves mostly correct pronunciation variant choice including glottal stop detection, algorithmically captures most of Czech assimilation logic and is both didactic and practical.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Václav Hanžl, Adléta Hanžlová. 2023-04-17. Prak: An automatic phonetic alignment tool for Czech. https://arxiv.org/abs/2304.08431
Cite the original work for its findings. Save a collection to share your selection of sources.