arXiv · 2505.07731
Spoken Language Understanding on Unseen Tasks With In-Context Learning
Abstract
Spoken language understanding (SLU) tasks involve diverse skills that probe the information extraction, classification and/or generation capabilities of models. In this setting, task-specific training data may not always be available. While traditional task-specific SLU models are unable to cater to such requirements, the speech-text large language models (LLMs) offer a promising alternative with emergent abilities. However, out of-the-box, our evaluations indicate that the zero/few-shot performance of prominent open-source speech-text LLMs on SLU tasks are not up to the mark. In this paper, we introduce a novel approach to robust task-agnostic fine-tuning using randomized class labels. With this proposed fine-tuning, we illustrate that the performance of the speech-text LLMs on an unseen task is significantly improved over standard approaches. Critically, the proposed approach avoids the requirement of task-specific data annotations for enabling new tasks in speech-text LLMs.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Neeraj Agrawal, Sriram Ganapathy. 2025-05-12. Spoken Language Understanding on Unseen Tasks With In-Context Learning. https://doi.org/10.21437/interspeech.2025-1467
Cite the original work for its findings. Save a collection to share your selection of sources.