A linear-time algorithm for computing the multinomial stochastic complexity

Petri Kontkanen,Petri Myllymäki

doi:10.1016/j.ipl.2007.04.003

Abstract

The minimum description length (MDL) principle is a theoretically well-founded, general framework for performing model class selection and other types of statistical inference. This framework can be applied for tasks such as data clustering, density estimation and image denoising. The MDL principle is formalized via the so-called normalized maximum likelihood (NML) distribution, which has several desirable theoretical properties. The codelength of a given sample of data under the NML distribution is called the stochastic complexity, which is the basis for MDL model class selection. Unfortunately, in the case of discrete data, straightforward computation of the stochastic complexity requires exponential time with respect to the sample size, since the definition involves an exponential sum over all the possible data samples of a fixed size. As a main contribution of this paper, we derive an elegant recursion formula which allows efficient computation of the stochastic complexity in the case of n observations of a single multinomial random variable with K values. The time complexity of the new method is O ( n + K ) as opposed to O ( n log n log K ) obtained with the previous results.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

A linear-time algorithm for computing the multinomial stochastic complexity

Abstract

Talk to us

Similar Papers

More From: Information Processing Letters

Lead the way for us

Journal: Information Processing Letters	Publication Date: Apr 20, 2007
Citations: 87

Similar Papers

Improved spatially adaptive MDL denoising of images using normalized maximum likelihood density
Srinivasan Meena ... S Annadurai
Image and Vision Computing | VOL. 26
Srinivasan Meena, et. al.Srinivasan Meena ... S Annadurai
01 May 2008
Image and Vision Computing | VOL. 26

The Minimum Description Length Principle
Peter D Grünwald
-
Peter D GrünwaldPeter D Grünwald
23 Mar 2007
23 Mar 2007

Enhanced minimum description length preprocessing of time series trajectories
Gajanan Gawde ... Jyoti Pawar
-
Gajanan Gawde, et. al.Gajanan Gawde ... Jyoti Pawar
01 Mar 2017
01 Mar 2017

Minimum Description Length Principle in Supervised Learning With Application to Lasso
Masanori Kawakita ... Jun'Ichi Takeuchi
IEEE Transactions on Information Theory | VOL. 66
Masanori Kawakita, et. al.Masanori Kawakita ... Jun'Ichi Takeuchi
01 Jul 2020
IEEE Transactions on Information Theory | VOL. 66

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

A linear-time algorithm for computing the multinomial stochastic complexity

Abstract

Talk to us

Similar Papers

More From: Information Processing Letters