Statistical Problem Classes and Their Links to Information Theory |
| |
Authors: | Bertrand Clarke Jennifer Clarke Chi Wai Yu |
| |
Institution: | 1. Departments of Medicine and of Epidemiology and Public Health, Center for Computational Science , University of Miami , Miami , Florida , USA bclarke2@med.miami.edu;3. Department of Epidemiology and Public Health , University of Miami , Miami , Florida , USA;4. Department of Mathematics , Hong Kong University of Science and Technology , Hong Kong |
| |
Abstract: | We begin by recalling the tripartite division of statistical problems into three classes, M-closed, M-complete, and M-open and then reviewing the key ideas of introductory Shannon theory. Focusing on the related but distinct goals of model selection and prediction, we argue that different techniques for these two goals are appropriate for the three different problem classes. For M-closed problems we give relative entropy justification that the Bayes information criterion (BIC) is appropriate for model selection and that the Bayes model average is information optimal for prediction. For M-complete problems, we discuss the principle of maximum entropy and a way to use the rate distortion function to bypass the inaccessibility of the true distribution. For prediction in the M-complete class, there is little work done on information based model averaging so we discuss the Akaike information criterion (AIC) and its properties and variants. For the M-open class, we argue that essentially only predictive criteria are suitable. Thus, as an analog to model selection, we present the key ideas of prediction along a string under a codelength criterion and propose a general form of this criterion. Since little work appears to have been done on information methods for general prediction in the M-open class of problems, we mention the field of information theoretic learning in certain general function spaces. |
| |
Keywords: | Bayesian Codelength Entropy Information theory M-closed M-complete M-open Mutual information Model selection Prediction Relative entropy Rate distortion |
|
|