Gaussian Component based Index for GMMs

www.lmu.de | UB | Blättern | Hilfe

Zur erweiterten Suche

English

Zur erweiterten Suche

Zhou, Linfei; Wackersreuther, Bianca; Fiedler, Frank; Plant, Claudia und Böhm, Christian (2016): Gaussian Component based Index for GMMs. In: 2016 IEEE 16th international Conference On Data Mining (ICDM): S. 1365-1370

Volltext auf 'Open Access LMU' nicht verfügbar.

DOI: 10.1109/ICDM.2016.0187

Abstract

Efficient similarity search for uncertain data is a challenging task in many modern data mining applications like image retrieval, speaker recognition and stock market analysis. A common way to model the uncertainty of data objects is using probability density functions in the form of Gaussian Mixture Models (GMMs), which have an ability to approximate arbitrary distribution. However, due to the possible unequal length of mixture models, the use of existing index techniques has serious problems for the objects modeled by GMMs. Either the techniques cannot handle GMMs or they have too many limitations. Hence, we propose a dynamic index structure, Gaussian Component based Index (GCI), for GMMs. GCI decomposes GMMs into the single, pairs, or n-lets of Gaussian components, stores these components into well studied index trees such as U-tree and Gauss-Tree, and refines the corresponding GMMs in a conservative but tight way. GCI supports both k-mostlikely queries and probability threshold queries by means of Matching Probability. Extensive experimental evaluations of GCI demonstrate a considerable speed-up of similarity search on both synthetic and real-world data sets.

Dokumententyp:	Zeitschriftenartikel
Fakultät:	Mathematik, Informatik und Statistik > Informatik
Themengebiete:	000 Informatik, Informationswissenschaft, allgemeine Werke > 004 Informatik
ISSN:	1550-4786
Sprache:	Englisch
Dokumenten ID:	47374
Datum der Veröffentlichung auf Open Access LMU:	27. Apr. 2018 08:12
Letzte Änderungen:	13. Aug. 2024 12:54

Dokument bearbeiten