arXiv · 2601.05289
A universal vision transformer for fast calorimeter simulations
Abstract
The high-dimensional complex nature of detectors makes fast calorimeter simulations a prime application for modern generative machine learning. Vision transformers (ViTs) can emulate the Geant4 response with unmatched accuracy and are not limited to regular geometries. Starting from the CaloDREAM architecture, we demonstrate the robustness and scalability of ViTs on regular and irregular geometries, and multiple detectors. Our results show that ViTs generate electromagnetic and hadronic showers with minimal deviations from Geant4 in multiple evaluation metrics, while maintaining the generation time in the $\mathcal{O}(10-100)$ ms on a single GPU. Furthermore, we show that pretraining on a large dataset and fine-tuning on the target geometry leads to reduced training costs and higher data efficiency, or altogether improves the fidelity of generated showers.
Explore related subjects
Keep this discovery
Luigi Favaro, Andrea Giammanco, Claudius Krause. 2026-01-07. A universal vision transformer for fast calorimeter simulations. https://doi.org/10.1088/2632-2153%2Fae7179
Cite the original work for its findings. Save a collection to share your selection of sources.