cs.AISep 28, 2026

HyperZip: Efficient Data Compression through Personalized Diffusion LLMs with Hypernetworks

Authors: Thai Nguyen, Khang Tran, NhatHai Phan

Organizations: New Jersey Institute of Technology

Abstract

Large language models (LLMs) have shown strong potential for lossless data compression, but existing approaches are constrained by the high computational cost and low throughput of autoregressive decoding. We propose HyperZip, an efficient and scalable LLM-based compression framework that leverages diffusion-based LLMs (dLLMs) with Multi-Token Prediction (MTP) to accelerate LLM-based data compression processes. We identify a trade-off in diffusion-based compression, where increasing decoding throughput degrades the compression rate. To mitigate this trade-off, HyperZip employs a hypernetwork to generate data-specific updates from a context representation, adapting the dLLM to the target data without costly fine-tuning, resulting in a low compression rate and high throughput. Extensive experiments show that HyperZip achieves a superior trade-off between compression rate and speed compared with state-of-the-art baselines.

Figures & tables

Appendix figures & tables4 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Diffuse to Compress: Leveraging Diffusion LMs for Lossless Compression

    Aug 4, 2026Angelo Nardone, Paolo FerraginaLarge Language Model CompressionData Compression Methods

  2. ZipCCL: Efficient Lossless Data Compression of Communication Collectives for Accelerating LLM Training

    Apr 30, 2026Wenxiang Lin, Xinglin Pan, Ruibo Fan +2Large Language Model CompressionLarge Language Model Training

  3. LLM-based Source Code Compression via Thresholded Symbol Ranking

    Jul 27, 2026Angelo Nardone, Paolo FerraginaLarge Language Model CompressionTask-Aware Compression