Mixed Membership Stochastic Blockmodels

doi:10.5555/1390681.1442798

Open AccessJournal ArticleDOI

Mixed Membership Stochastic Blockmodels

Edoardo M. Airoldi, +3 more

- 01 Jun 2008 -

Journal of Machine Learning Research

- Vol. 9, Iss: 65, pp 1981-2014

TLDR

In this article, the authors introduce a class of variance allocation models for pairwise measurements, called mixed membership stochastic blockmodels, which combine global parameters that instantiate dense patches of connectivity (blockmodel) with local parameters (mixed membership), and develop a general variational inference algorithm for fast approximate posterior inference.

Abstract:

Consider data consisting of pairwise measurements, such as presence or absence of links between pairs of objects. These data arise, for instance, in the analysis of protein interactions and gene regulatory networks, collections of author-recipient email, and social networks. Analyzing pairwise measurements with probabilistic models requires special assumptions, since the usual independence or exchangeability assumptions no longer hold. Here we introduce a class of variance allocation models for pairwise measurements: mixed membership stochastic blockmodels. These models combine global parameters that instantiate dense patches of connectivity (blockmodel) with local parameters that instantiate node-specific variability in the connections (mixed membership). We develop a general variational inference algorithm for fast approximate posterior inference. We demonstrate the advantages of mixed membership stochastic blockmodels with applications to social networks and protein interaction networks.

Citations

PDF

Open Access

More filters

Proceedings ArticleDOI

Selecting the Best Solvers: Toward Community Based Crowdsourcing for Disaster Management

Zhiyong Yu, +3 more

TL;DR: This paper designed a framework for community based crowd sourcing, i.e., task takers are from an existing community or will easily form a new community, and a size-specified community creation method using multiple social contexts is proposed.

...read moreread less

Journal ArticleDOI

Improvements on SCORE, Especially for Weak Signals

Jiashun Jin, +2 more

TL;DR: It is shown that in a broad class of network settings where the authors allow for weak signals, severe degree heterogeneity, and a wide range of network sparsity, SCORE achieves prefect clustering and has the so-called “exponential rate” in Hamming clustering errors.

...read moreread less

Journal ArticleDOI

On Ising models and algorithms for the construction of symptom networks in psychopathological research.

Michael J. Brusco, +4 more

- 07 Oct 2019 -

Psychological Methods

TL;DR: This article provides a careful assessment of the conditions that underlie the Ising model as well as specific limitations associated with the eLasso estimation algorithm, which leads to serious concerns regarding the implementation ofeLasso in psychopathological research.

...read moreread less

Proceedings ArticleDOI

Most large topic models are approximately separable

Weicong Ding, +2 more

TL;DR: It is proved that when the columns of the topic matrix are independently sampled from a Dirichlet distribution, the resulting topic matrix will be approximately separable with probability tending to one as the number of rows (vocabulary size) scales to infinity sufficiently faster than thenumber of columns (topics).

...read moreread less

Patent

Systems and methods for genomic pattern analysis

Vladimir Semenyuk

TL;DR: In this article, the authors propose a method for analyzing sequence data in which a large amount and variety of reference data are efficiently modeled as a reference graph, such as a directed acyclic graph (DAG).

...read moreread less

Collapse

References

PDF

Open Access

More filters

Journal ArticleDOI

Maximum likelihood from incomplete data via the EM algorithm

Arthur P. Dempster, +2 more

- 01 Sep 1977 -

Journal of the royal statistical society...

Journal ArticleDOI

Latent dirichlet allocation

David M. Blei, +2 more

- 01 Mar 2003 -

Journal of Machine Learning Research

TL;DR: This work proposes a generative model for text and other collections of discrete data that generalizes or improves on several previous models including naive Bayes/unigram, mixture of unigrams, and Hofmann's aspect model.

...read moreread less

Journal ArticleDOI

Finding scientific topics

Thomas L. Griffiths, +1 more

- 06 Apr 2004 -

Proceedings of the National Academy of S...

TL;DR: A generative model for documents is described, introduced by Blei, Ng, and Jordan, and a Markov chain Monte Carlo algorithm is presented for inference in this model, which is used to analyze abstracts from PNAS by using Bayesian model selection to establish the number of topics.

...read moreread less

Journal ArticleDOI

Functional organization of the yeast proteome by systematic analysis of protein complexes

Anne-Claude Gavin, +37 more

- 10 Jan 2002 -

Nature

TL;DR: The analysis provides an outline of the eukaryotic proteome as a network of protein complexes at a level of organization beyond binary interactions, which contains fundamental biological information and offers the context for a more reasoned and informed approach to drug discovery.

...read moreread less