arXiv · 2601.08535
Distribution Estimation with Side Information
Abstract
We consider the classical problem of discrete distribution estimation using i.i.d. samples in a novel scenario where additional side information is available on the distribution. In large alphabet datasets such as text corpora, such side information arises naturally through word semantics/similarities that can be inferred by closeness of vector word embeddings, for instance. We consider two specific models for side information--a local model where the unknown distribution is in the neighborhood of a known distribution, and a partial ordering model where the alphabet is partitioned into known higher and lower probability sets. In both models, we theoretically characterize the improvement in a suitable squared-error risk because of the available side information. Simulations over natural language and synthetic data illustrate these gains.
Explore related subjects
Keep this discovery
Haricharan Balasundaram, Andrew Thangaraj. 2026-01-13. Distribution Estimation with Side Information. https://arxiv.org/abs/2601.08535
Cite the original work for its findings. Save a collection to share your selection of sources.