Skip to main navigation Skip to search Skip to main content

Hyper-Compression: Model Compression via Hyperfunction

  • Feng-Lei Fan (Co-first Author)
  • , Juntong Fan (Co-first Author)
  • , Dayang Wang (Co-first Author)
  • , Jingbo Zhang
  • , Zelin Dong
  • , Shijun Zhang
  • , Ge Wang
  • , Tieyong Zeng*
  • *Corresponding author for this work

Research output: Journal Publications and ReviewsRGC 21 - Publication in refereed journalpeer-review

Abstract

The rapid growth of large models’ size has far outpaced that of computing resources. To bridge this gap, encouraged by the parsimonious relationship between genotype and phenotype in the brain’s growth and development, we propose the so-called Hyper-Compression that turns the model compression into the issue of parameter representation via a hyperfunction. Specifically, it is known that the trajectory of some low-dimensional dynamic systems can fill the high-dimensional space eventually. Thus, Hyper-Compression, using these dynamic systems as the hyperfunctions, represents the parameters of the target network by their corresponding composition number or trajectory length. This suggests a novel mechanism for model compression, substantially different from the existing pruning, quantization, distillation, and decomposition. Along this direction, we methodologically identify a suitable dynamic system with the irrational winding as the hyperfunction and theoretically derive its associated error bound. Next, guided by our theoretical insights, we propose several engineering twists to make the Hyper-Compression pragmatic and effective. Lastly, systematic and comprehensive experiments on NLP models such as LLaMA and Qwen series and vision models confirm that Hyper-Compression enjoys the following PNAS merits: 1) Preferable compression ratio; 2) No post-hoc retraining; 3) Affordable inference time; and 4) Short compression time. It compresses LLaMA2-7B in an hour and achieves close-to-int4-quantization performance, without retraining and with a performance drop of less than 1%. © 1979-2012 IEEE.
Original languageEnglish
Number of pages18
JournalIEEE Transactions on Pattern Analysis and Machine Intelligence
DOIs
Publication statusOnline published - 13 Mar 2026

Research Keywords

  • Dynamic System
  • Hyper-Compression
  • Large Models
  • Model Compression

Fingerprint

Dive into the research topics of 'Hyper-Compression: Model Compression via Hyperfunction'. Together they form a unique fingerprint.

Cite this