SearcharxivSearch

arXiv subjects

Wen-Yang Lu

Publications and source records attributed to Wen-Yang Lu.

2 recordsLinked to original sources

Reduced-complexity Adaptive Loop Filtering via Input-dependent Graph Filters

Adaptive Loop Filtering is an important tool for suppressing compression artifacts in modern video codecs. In the enhanced compression model (ECM), a software test model used for experimenting with video coding tools beyond Versatile Video Coding, fixed filters are trained offline and achieve high signal adaptivity via a fine-grained gradient-based classifier, resulting in a large number of fixed filters that introduce redundancy and increased implementation complexity. Reducing this redundancy without compromising artifact suppression, therefore, remains a key challenge. This paper proposes an alternative graph-based fixed-filtering framework for adaptive loop filtering. By using a graph to encode pixel-intensity relationships, our approach captures local structural information more effectively than gradient-based classification alone. Fixed filters are learned as polynomial graph filters, enabling structurally similar local patterns to share common filtering behavior. Experimental results demonstrate that the proposed approach achieves a comparable performance to the ECM baseline while reducing the number of required filters by an order of magnitude.

eess.IV

Adaptive Online Learning of Separable Path Graph Transforms for Intra-prediction

Current video coding standards, including H.264/AVC, HEVC, and VVC, employ discrete cosine transform (DCT), discrete sine transform (DST), and secondary to Karhunen-Loeve transforms (KLTs) decorrelate the intra-prediction residuals. However, the efficiency of these transforms in decorrelation can be limited when the signal has a non-smooth and non-periodic structure, such as those occurring in textures with intricate patterns. This paper introduces a novel adaptive separable path graph-based transform (GBT) that can provide better decorrelation than the DCT for intra-predicted texture data. The proposed GBT is learned in an online scenario with sequential K-means clustering, which groups similar blocks during encoding and decoding to adaptively learn the GBT for the current block from previously reconstructed areas with similar characteristics. A signaling overhead is added to the bitstream of each coding block to indicate the usage of the proposed graph-based transform. We assess the performance of this method combined with H.264/AVC intra-coding tools and demonstrate that it can significantly outperform H.264/AVC DCT for intra-predicted texture data.

eess.IV