arXiv · 2409.12433
A High-Throughput Hardware Accelerator for Lempel-Ziv 4 Compression Algorithm
Abstract
This paper delves into recent hardware implementations of the Lempel-Ziv 4 (LZ4) algorithm, highlighting two key factors that limit the throughput of single-kernel compressors. Firstly, the actual parallelism exhibited in single-kernel designs falls short of the theoretical potential. Secondly, the clock frequency is constrained due to the presence of the feedback loops. To tackle these challenges, we propose a novel scheme that restricts each parallelization window to a single match, thus elevating the level of actual parallelism. Furthermore, by restricting the maximum match length, we eliminate the feedback loops within the architecture, enabling a significant boost in throughput. Finally, we present a high-speed hardware architecture. The implementation results demonstrate that the proposed architecture achieves a throughput of up to 16.10 Gb/s, exhibiting a 2.648x improvement over the start-of-the-art. The new design only results in an acceptable compression ratio reduction ranging from 4.93% to 11.68% with various numbers of hash table entries, compared to the LZ4 compression ratio achieved by official software implementations disclosed on GitHub.
Explore related subjects
Keep this discovery
Tao Chen, Suwen Song, Zhongfeng Wang. 2024-09-19. A High-Throughput Hardware Accelerator for Lempel-Ziv 4 Compression Algorithm. https://arxiv.org/abs/2409.12433
Cite the original work for its findings. Save a collection to share your selection of sources.