An evaluation of data stream clustering algorithms

Data stream clustering is a hot research area due to the abundance of data streams collected nowadays and the need for understanding and acting upon such sort of data. Unsupervised learning (clustering) comprises one of the most popular data mining tasks for gaining insights into the data. Clusterin...

Celý popis

Uložené v:
Podrobná bibliografia
Vydané v:Statistical analysis and data mining Ročník 11; číslo 4; s. 167 - 187
Hlavní autori: Mansalis, Stratos, Ntoutsi, Eirini, Pelekis, Nikos, Theodoridis, Yannis
Médium: Journal Article
Jazyk:English
Vydavateľské údaje: Hoboken Wiley Subscription Services, Inc., A Wiley Company 01.08.2018
Wiley Subscription Services, Inc
Predmet:
ISSN:1932-1864, 1932-1872
On-line prístup:Získať plný text
Tagy: Pridať tag
Žiadne tagy, Buďte prvý, kto otaguje tento záznam!
Popis
Shrnutí:Data stream clustering is a hot research area due to the abundance of data streams collected nowadays and the need for understanding and acting upon such sort of data. Unsupervised learning (clustering) comprises one of the most popular data mining tasks for gaining insights into the data. Clustering is a challenging task, while clustering over data streams involves additional challenges such as the single pass constraint over the raw data and the need for fast response. Moreover, dealing with an infinite and fast changing data stream implies that the clustering model extracted upon such sort of data is also subject to evolution over time. Several stream clustering surveys exist already in the literature; however, they focus on a theoretical presentation of the surveyed algorithms. On the contrary, in this paper, we survey the state‐of‐the‐art stream clustering algorithms and we evaluate their performance in different data sets and for different parameter settings.
Bibliografia:ObjectType-Article-1
SourceType-Scholarly Journals-1
ObjectType-Feature-2
content type line 14
ISSN:1932-1864
1932-1872
DOI:10.1002/sam.11380