PASCAL - Pattern Analysis, Statistical Modelling and Computational Learning

Data streaming with Affinity propagation
Xiangliang Zhang, Cyril Furtlehner and Michele Sebag
In: ECML2008, 15-19 September 2008, Antwerp, Belgium.


This paper proposed StrAP (Streaming AP), extending Affinity Propagation (AP) to data steaming. AP, a new clustering algorithm, extracts the data items, or exemplars, that best represent the dataset using a message passing method. Several steps are made to build StrAP. The first one (Weighted AP) extends AP to weighted items with no loss of generality. The second one (Hierarchical WAP) is concerned with reducing the quadratic AP complexity, by applying AP on data subsets and further applying Weighted AP on the exemplars extracted from all subsets. Finally StrAP extends Hierarchical WAP to deal with changes in the data distribution. Experiments on artificial datasets, on the Intrusion Detection benchmark (KDD99) and on a real-world problem, clustering the stream of jobs submitted to the EGEE grid system, provide a comparative validation of the approach.

PDF (paper) - Requires Adobe Acrobat Reader or other PDF viewer.
PDF (Slides) - Requires Adobe Acrobat Reader or other PDF viewer.
EPrint Type:Conference or Workshop Item (Paper)
Project Keyword:Project Keyword UNSPECIFIED
Subjects:Learning/Statistics & Optimisation
ID Code:4492
Deposited By:Xiangliang Zhang
Deposited On:13 March 2009