arXiv · 2207.14473
Benchmarking Azerbaijani Neural Machine Translation
Abstract
Little research has been done on Neural Machine Translation (NMT) for Azerbaijani. In this paper, we benchmark the performance of Azerbaijani-English NMT systems on a range of techniques and datasets. We evaluate which segmentation techniques work best on Azerbaijani translation and benchmark the performance of Azerbaijani NMT models across several domains of text. Our results show that while Unigram segmentation improves NMT performance and Azerbaijani translation models scale better with dataset quality than quantity, cross-domain generalization remains a challenge
Explore related subjects
Keep this discovery
Chih-Chen Chen, William Chen. 2022-07-29. Benchmarking Azerbaijani Neural Machine Translation. https://arxiv.org/abs/2207.14473
Cite the original work for its findings. Save a collection to share your selection of sources.