arXiv · 2307.08189
Mini-Giants: "Small" Language Models and Open Source Win-Win
Abstract
ChatGPT is phenomenal. However, it is prohibitively expensive to train and refine such giant models. Fortunately, small language models are flourishing and becoming more and more competent. We call them "mini-giants". We argue that open source community like Kaggle and mini-giants will win-win in many ways, technically, ethically and socially. In this article, we present a brief yet rich background, discuss how to attain small language models, present a comparative study of small language models and a brief discussion of evaluation methods, discuss the application scenarios where small language models are most needed in the real world, and conclude with discussion and outlook.
Explore related subjects
Keep this discovery
Zhengping Zhou, Lezhi Li, Xinxi Chen, Andy Li. 2023-07-17. Mini-Giants: "Small" Language Models and Open Source Win-Win. https://arxiv.org/abs/2307.08189
Cite the original work for its findings. Save a collection to share your selection of sources.