arXiv · 2108.11463
With One Voice: Composing a Travel Voice Assistant from Re-purposed Models
Abstract
Voice assistants provide users a new way of interacting with digital products, allowing them to retrieve information and complete tasks with an increased sense of control and flexibility. Such products are comprised of several machine learning models, like Speech-to-Text transcription, Named Entity Recognition and Resolution, and Text Classification. Building a voice assistant from scratch takes the prolonged efforts of several teams constructing numerous models and orchestrating between components. Alternatives such as using third-party vendors or re-purposing existing models may be considered to shorten time-to-market and development costs. However, each option has its benefits and drawbacks. We present key insights from building a voice search assistant for Booking.com search and recommendation system. Our paper compares the achieved performance and development efforts in dedicated tailor-made solutions against existing re-purposed models. We share and discuss our data-driven decisions about implementation trade-offs and their estimated outcomes in hindsight, showing that a fully functional machine learning product can be built from existing models.
Explore related subjects
Keep this discovery
Shachaf Poran, Gil Amsalem, Amit Beka, Dmitri Goldenberg. 2021-08-04. With One Voice: Composing a Travel Voice Assistant from Re-purposed Models. https://arxiv.org/abs/2108.11463
Cite the original work for its findings. Save a collection to share your selection of sources.