arXiv · 2509.09701
Optimal Multi-Task Learning at Regularization Horizon for Speech Translation Task
Abstract
End-to-end speech-to-text translation typically suffers from the scarcity of paired speech-text data. One way to overcome this shortcoming is to utilize the bitext data from the Machine Translation (MT) task and perform Multi-Task Learning (MTL). In this paper, we formulate MTL from a regularization perspective and explore how sequences can be regularized within and across modalities. By thoroughly investigating the effect of consistency regularization (different modality) and R-drop (same modality), we show how they respectively contribute to the total regularization. We also demonstrate that the coefficient of MT loss serves as another source of regularization in the MTL setting. With these three sources of regularization, we introduce the optimal regularization contour in the high-dimensional space, called the regularization horizon. Experiments show that tuning the hyperparameters within the regularization horizon achieves near state-of-the-art performance on the MuST-C dataset.
Explore related subjects
Keep this discovery
JungHo Jung, Junhyun Lee. 2025-09-04. Optimal Multi-Task Learning at Regularization Horizon for Speech Translation Task. https://arxiv.org/abs/2509.09701
Cite the original work for its findings. Save a collection to share your selection of sources.