arXiv · 2609.36357
HyperZip: Efficient Data Compression through Personalized Diffusion LLMs with Hypernetworks
Abstract
Large language models (LLMs) have shown strong potential for lossless data compression, but existing approaches are constrained by the high computational cost and low throughput of autoregressive decoding. We propose HyperZip, an efficient and scalable LLM-based compression framework that leverages diffusion-based LLMs (dLLMs) with Multi-Token Prediction (MTP) to accelerate LLM-based data compression processes. We identify a trade-off in diffusion-based compression, where increasing decoding throughput degrades the compression rate. To mitigate this trade-off, HyperZip employs a hypernetwork to generate data-specific updates from a context representation, adapting the dLLM to the target data without costly fine-tuning, resulting in a low compression rate and high throughput. Extensive experiments show that HyperZip achieves a superior trade-off between compression rate and speed compared with state-of-the-art baselines.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Thai Nguyen, Khang Tran, NhatHai Phan. 2026-09-28. HyperZip: Efficient Data Compression through Personalized Diffusion LLMs with Hypernetworks. https://arxiv.org/abs/2609.36357
Cite the original work for its findings. Save a collection to share your selection of sources.