Abstract
Deep-learning-based compressor has received interests recently due to much improved compression ratio. However, modern approaches suffer from long execution time. To ease this problem, this paper targets on cutting down the execution time of deep-learning-based compressors. Building history-dependencies sequentially (e.g., recurrent neural networks) is responsible for long inference latency. Instead, we introduce transformer into deep learning compressors to build history-dependencies in parallel. However, existing transformer is too heavy in computation and incompatible to compression tasks.
This paper proposes a fast general-purpose lossless compressor, TRACE, by designing a compression-friendly structure based on a single-layer transformer. We first design a new metric to advise the selection part of compression model structures. Byte-grouping and Shared-ffn schemes are further proposed to fully utilize the capacity of the single-layer transformer. These features allow TRACE to achieve competitive compression ratio and a much faster speed. In addition, we further accelerate the compression procedure by designing a controller to reduce the parameter updating overhead. Experiments show that TRACE achieves an overall 3x speedup while keeps a comparable compression ratio to the state-of-the-art compressors. The source code for TRACE and links to the datasets are available at https://github.com/mynotwo/A-Fast-Transformer-based-General-Purpose-LosslessCompressor.
This paper proposes a fast general-purpose lossless compressor, TRACE, by designing a compression-friendly structure based on a single-layer transformer. We first design a new metric to advise the selection part of compression model structures. Byte-grouping and Shared-ffn schemes are further proposed to fully utilize the capacity of the single-layer transformer. These features allow TRACE to achieve competitive compression ratio and a much faster speed. In addition, we further accelerate the compression procedure by designing a controller to reduce the parameter updating overhead. Experiments show that TRACE achieves an overall 3x speedup while keeps a comparable compression ratio to the state-of-the-art compressors. The source code for TRACE and links to the datasets are available at https://github.com/mynotwo/A-Fast-Transformer-based-General-Purpose-LosslessCompressor.
| Original language | English |
|---|---|
| Title of host publication | WWW '22 |
| Subtitle of host publication | Proceedings of the ACM Web Conference 2022 |
| Editors | Frédérique Laforest, Raphaël Troncy, Elena Simperl, Deepak Agarwal, Aristides Gionis, Ivan Herman, Lionel Médini |
| Publisher | Association for Computing Machinery |
| Pages | 1829-1838 |
| ISBN (Print) | 978-1-4503-9096-5 |
| DOIs | |
| Publication status | Published - Apr 2022 |
| Event | 31st ACM Web Conference (WWW 2022) - Virtual, Lyon, France Duration: 25 Apr 2022 → 29 Apr 2022 |
Publication series
| Name | WWW - Proceedings of the ACM Web Conference |
|---|
Conference
| Conference | 31st ACM Web Conference (WWW 2022) |
|---|---|
| Place | France |
| City | Lyon |
| Period | 25/04/22 → 29/04/22 |
Research Keywords
- byte stream
- computational efficient model
- general-purpose compressor
- lossless data compression
- neural networks
- transformer
Fingerprint
Dive into the research topics of 'TRACE: A Fast Transformer-based General-Purpose Lossless Compressor'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver