arXiv · 2606.24737
VSANet: View-aware Sparse Attention Network for Light Field Image Denoising
Abstract
Light field (LF) image denoising is challenging due to the high-dimensional structure of LF data. While noise is independent across sub-aperture images, scene content exhibits strong cross-view correlations. We introduce VSANet, a view-aware sparse attention network for LF denoising. Specifically, we propose a view-aware sparse attention (VSA) block that represents the 4D LF feature map as a unified spatial-angular token space and performs cross-view aggregation via locality-sensitive hashing-based sparse attention. This enables global feature interactions with linear complexity, effectively exploiting LF correlations across views and spatial locations. In addition, we design a feature refinement (FR) block to emphasize informative features in spatial, angular, and epipolar subspaces. The VSA and FR blocks are integrated within a sequential attention refinement module, forming the core of VSANet. Experiments demonstrate VSANet outperforms stateof-the-art LF denoising methods.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Gargi Panda, Soumitra Kundu, Saumik Bhattacharya, Aurobinda Routray. 2026-06-23. VSANet: View-aware Sparse Attention Network for Light Field Image Denoising. https://arxiv.org/abs/2606.24737
Cite the original work for its findings. Save a collection to share your selection of sources.