Asymptotically minimax regret by Bayes mixtures

📅 1998-08-16
🏛️ Proceedings. 1998 IEEE International Symposium on Information Theory (Cat. No.98CH36252)
📈 Citations: 46
Influential: 4
📄 PDF

career value

230K/year
🤖 AI Summary
This paper investigates the regret (redundancy) of sequential prediction, data compression, and gambling relative to smooth parametric families—specifically exponential families, general smooth parametric models, and Markov sources. It analyzes the asymptotic regret performance of Bayesian mixture distributions, particularly Jeffreys-prior-based variants, under maximum-likelihood estimation. The key contribution is the first rigorous proof that such Jeffreys-type mixture priors achieve asymptotically minimax regret over all these model classes, with redundancy converging at the optimal rate of $O(1/n)$. This rate matches the information-theoretic lower bound dictated by the Shtarkov normalized constant. By unifying tools from information geometry, asymptotic statistics, and normalized maximum-likelihood theory, the work establishes a fundamental connection between Bayesian mixtures and information-theoretic optimality. It thus provides a unified, theoretically grounded guarantee of optimality for universal coding and prediction.

Technology Category

Application Category

📝 Abstract
We study the problem of data compression, gambling and prediction of a sequence x/sup n/ = x/sub 1/x/sub 2/...x/sub n/ from a certain alphabet X, in terms of regret (Shtarkov 1988) and redundancy with respect to a general exponential family, a general smooth family, and also Markov sources. In particular, we show that variants of Jeffreys mixture asymptotically achieve their minimax values.
Problem

Research questions and friction points this paper is trying to address.

Evaluating regret of Bayes mixtures vs maximum likelihood
Achieving asymptotically minimax regret with Jeffreys prior variants
Extending minimax regret to non-exponential families via modifications
Innovation

Methods, ideas, or system contributions that make the work stand out.

Bayes mixtures achieve minimax regret
Modified Jeffreys prior for non-exponential families
Local exponential tilting enlarges family
🔎 Similar Papers
2024-10-02International Conference on Machine LearningCitations: 1