Gaussian Component Based Index for GMMs

Linfei Zhou,Bianca Wackersreuther,Frank Fiedler,Christian Bohm,Claudia Plant

doi:10.1109/icdm.2016.0187

Abstract

Efficient similarity search for uncertain data is a challenging task in many modern data mining applications like image retrieval, speaker recognition and stock market analysis. A common way to model the uncertainty of data objects is using probability density functions in the form of Gaussian Mixture Models (GMMs), which have an ability to approximate arbitrary distribution. However, due to the possible unequal length of mixture models, the use of existing index techniques has serious problems for the objects modeled by GMMs. Either the techniques cannot handle GMMs or they have too many limitations. Hence, we propose a dynamic index structure, Gaussian Component based Index (GCI), for GMMs. GCI decomposes GMMs into the single, pairs, or n-lets of Gaussian components, stores these components into well studied index trees such as U-tree and Gauss-Tree, and refines the corresponding GMMs in a conservative but tight way. GCI supports both k-most-likely queries and probability threshold queries by means of Matching Probability. Extensive experimental evaluations of GCI demonstrate a considerable speed-up of similarity search on both synthetic and real-world data sets.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Gaussian Component Based Index for GMMs

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Novel Indexing Strategy and Similarity Measures for Gaussian Mixture Models
Linfei Zhou ... Christian Böhm
-
Linfei Zhou, et. al.Linfei Zhou ... Christian Böhm
01 Jan 2017
01 Jan 2017

Searching Uncertain Data Represented by Non-axis Parallel Gaussian Mixture Models
Katrin Haegler ... Christian Bohm
-
Katrin Haegler, et. al.Katrin Haegler ... Christian Bohm
01 Apr 2012
01 Apr 2012

Measure of Similarity between GMMs Based on Autoencoder-Generated Gaussian Component Representations
Vladimir Kalušev ... Marko Janev
Axioms | VOL. 12
Vladimir Kalušev, et. al.Vladimir Kalušev ... Marko Janev
30 May 2023
Axioms | VOL. 12

Querying Objects Modeled by Arbitrary Probability Distributions
Christian Böhm ... Peter Kunath
-
Christian Böhm, et. al.Christian Böhm ... Peter Kunath
16 Jul 2007
16 Jul 2007

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Gaussian Component Based Index for GMMs

Abstract

Talk to us

Similar Papers