The Annals of Applied Statistics

Strategies for online inference of model-based clustering in large and growing networks

Hugo Zanghi, Franck Picard, Vincent Miele, and Christophe Ambroise

Open access


In this paper we adapt online estimation strategies to perform model-based clustering on large networks. Our work focuses on two algorithms, the first based on the SAEM algorithm, and the second on variational methods. These two strategies are compared with existing approaches on simulated and real data. We use the method to decipher the connexion structure of the political websphere during the US political campaign in 2008. We show that our online EM-based algorithms offer a good trade-off between precision and speed, when estimating parameters for mixture distributions in the context of random graphs.

Ann. Appl. Stat. Volume 4, Number 2 (2010), 687-714.

First available in Project Euclid: 3 August 2010

Graph clustering EM Algorithms online strategies web graph structure analysis


Zanghi, Hugo; Picard, Franck; Miele, Vincent; Ambroise, Christophe. Strategies for online inference of model-based clustering in large and growing networks. Ann. Appl. Stat. 4 (2010), no. 2, 687--714. doi:10.1214/10-AOAS359.

