arXiv · 2409.05258
Towards Automated Machine Learning Research
Abstract
This paper explores a top-down approach to automating incremental advances in machine learning research through component-level innovation, facilitated by Large Language Models (LLMs). Our framework systematically generates novel components, validates their feasibility, and evaluates their performance against existing baselines. A key distinction of this approach lies in how these novel components are generated. Unlike traditional AutoML and NAS methods, which often rely on a bottom-up combinatorial search over predefined, hardcoded base components, our method leverages the cross-domain knowledge embedded in LLMs to propose new components that may not be confined to any hard-coded predefined set. By incorporating a reward model to prioritize promising hypotheses, we aim to improve the efficiency of the hypothesis generation and evaluation process. We hope this approach offers a new avenue for exploration and contributes to the ongoing dialogue in the field.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Shervin Ardeshir. 2024-09-09. Towards Automated Machine Learning Research. https://arxiv.org/abs/2409.05258
Cite the original work for its findings. Save a collection to share your selection of sources.