An evaluation of data stream clustering algorithms

Data stream clustering is a hot research area due to the abundance of data streams collected nowadays and the need for understanding and acting upon such sort of data. Unsupervised learning (clustering) comprises one of the most popular data mining tasks for gaining insights into the data. Clusterin...

Full description

Saved in:
Bibliographic Details
Published in:Statistical analysis and data mining Vol. 11; no. 4; pp. 167 - 187
Main Authors: Mansalis, Stratos, Ntoutsi, Eirini, Pelekis, Nikos, Theodoridis, Yannis
Format: Journal Article
Language:English
Published: Hoboken Wiley Subscription Services, Inc., A Wiley Company 01.08.2018
Wiley Subscription Services, Inc
Subjects:
ISSN:1932-1864, 1932-1872
Online Access:Get full text
Tags: Add Tag
No Tags, Be the first to tag this record!
Description
Summary:Data stream clustering is a hot research area due to the abundance of data streams collected nowadays and the need for understanding and acting upon such sort of data. Unsupervised learning (clustering) comprises one of the most popular data mining tasks for gaining insights into the data. Clustering is a challenging task, while clustering over data streams involves additional challenges such as the single pass constraint over the raw data and the need for fast response. Moreover, dealing with an infinite and fast changing data stream implies that the clustering model extracted upon such sort of data is also subject to evolution over time. Several stream clustering surveys exist already in the literature; however, they focus on a theoretical presentation of the surveyed algorithms. On the contrary, in this paper, we survey the state‐of‐the‐art stream clustering algorithms and we evaluate their performance in different data sets and for different parameter settings.
Bibliography:ObjectType-Article-1
SourceType-Scholarly Journals-1
ObjectType-Feature-2
content type line 14
ISSN:1932-1864
1932-1872
DOI:10.1002/sam.11380