arXiv · 2505.17626
Leveraging Stochastic Depth Training for Adaptive Inference
Abstract
Dynamic DNN optimization techniques such as layer-skipping offer increased adaptability and efficiency gains but can lead to i) a larger memory footprint as in decision gates, ii) increased training complexity (e.g., with non-differentiable operations), and iii) less control over performance-quality trade-offs due to its inherent input-dependent execution. To approach these issues, we propose a simpler yet effective alternative for adaptive inference with a zero-overhead, single-model, and time-predictable inference. Central to our approach is the observation that models trained with Stochastic Depth -- a method for faster training of residual networks -- become more resilient to arbitrary layer-skipping at inference time. We propose a method to first select near Pareto-optimal skipping configurations from a stochastically-trained model to adapt the inference at runtime later. Compared to original ResNets, our method shows improvements of up to 2X in power efficiency at accuracy drops as low as 0.71%.
Explore related subjects
Keep this discovery
Guilherme Korol, Antonio Carlos Schneider Beck, Jeronimo Castrillon. 2025-05-23. Leveraging Stochastic Depth Training for Adaptive Inference. https://arxiv.org/abs/2505.17626
Cite the original work for its findings. Save a collection to share your selection of sources.