arXiv · 2508.17826
LLMulator: Generalizable Cost Modeling for Dataflow Accelerators with Input-Adaptive Control Flow
Abstract
Accurate and fast performance prediction for dataflow-based accelerators is vital for efficient hardware design and design space exploration, yet existing methods struggle to generalize across architectures, applications, and input-dependent control flows. We present LLMulator, a progressive numeric modeling framework leveraging the program semantic knowledge of pre-trained large language models (LLMs) for robust, hardware- and application-aware prediction. Our numeric model treats performance values as categorical token sequences, enabling range-agnostic estimates and confidence-aware predictions for unseen applications. To handle input-dependent control flows, we introduce a reinforcement learning-based dynamic calibration method, reducing cycle prediction error by 9.7% over static models and converging to 11.2% error after a few iterations. For cross-hardware generalization, we develop a progressive data augmentation strategy that generates diverse datasets covering multi-level dataflow structures, memory parameters, and loop mapping primitives, significantly boosting prediction accuracy across architectures and configurations.
Explore related subjects
Keep this discovery
Kaiyan Chang, Wenlong Zhu, Shengwen Liang, Huawei Li, Ying Wang. 2025-08-25. LLMulator: Generalizable Cost Modeling for Dataflow Accelerators with Input-Adaptive Control Flow. https://arxiv.org/abs/2508.17826
Cite the original work for its findings. Save a collection to share your selection of sources.