arXiv · 2605.16709
Covert Multi-bit LLM Watermarking: An Information Theory and Coding Approach
Abstract
We study the problem of multi-bit watermarking for non-autoregressive large language models (LLMs). We introduce an information-theoretic model inspired by diffusion language models (DLM), in which the encoder has limited non-causal access to token distributions within each token block. This formulation enables an information-theoretic characterization of the non-causal watermarking capacity, in which knowledge of LLM cover statistics is leveraged to enable a multi-bit covert embedding. We study the information-theoretic limits of the model by combining Gelfand--Pinsker and channel synthesis coding techniques and obtain an exact characterization of the capacity. The embedding strategy is further optimized across blocks using a constrained Markov decision process (CMDP) and we develop an explicit algorithm based on polar codes following the information-theoretic principles. We simulate the error performance on LLaDA, and provide empirical total variation (TV) analysis as a function of key randomness.
Explore related subjects
Keep this discovery
Sidong Guo, Tyler Kann, Teodora Baluta, Matthieu R. Bloch. 2026-09-02. Covert Multi-bit LLM Watermarking: An Information Theory and Coding Approach. https://arxiv.org/abs/2605.16709
Cite the original work for its findings. Save a collection to share your selection of sources.
Discover connections
Connections use source metadata and explicit phrase matches, not verified experimental comparisons.