arXiv · 2608.01343
DeVIT: Low-Power Vision Transformer Acceleration Using Delta Computation
Abstract
The emergence of transformer-based deep learning models has brought unprecedented performance across various domains, particularly in natural language processing and computer vision. However, deploying these models, especially on resource-constrained devices, poses significant challenges due to their high computational complexity and large memory size and bandwidth requirements. This complexity has led researchers to use low-bit model weights to reduce memory usage and improve efficiency. In addition to reducing processing and memory demands, quantization introduces another useful property: value locality, where the extremely large number of parameters are restricted to a limited range of values. To fully take advantage of this locality, this paper presents DeVIT, an acceleration method for vision transformers that leverages differential computation to enable multiplier-less matrix multiplication.
Explore related subjects
Keep this discovery
Reyhaneh Hosseinzadeh, Parham Zilouchian Moghaddam, Mehdi Modarressi. 2026-08-02. DeVIT: Low-Power Vision Transformer Acceleration Using Delta Computation. https://arxiv.org/abs/2608.01343
Cite the original work for its findings. Save a collection to share your selection of sources.