Bipartite Isoperimetric Graph Partitioning for Data Co-clustering

Show full item record

Title: Bipartite Isoperimetric Graph Partitioning for Data Co-clustering
Author: Rege, Manjeet; Dong, Ming; Fotouhi, Farshad
Abstract: Data co-clustering refers to the problem of simultaneous clustering of two data types. Typically, the data is stored in a contingency or co-occurrence matrix C where rows and columns of the matrix represent the data types to be co-clustered. An entry Cij of the matrix signifies the relation between the data type represented by row i and column j. Co-clustering is the problem of deriving sub-matrices from the larger data matrix by simultaneously clustering rows and columns of the data matrix. In this paper, we present a novel graph theoretic approach to data co-clustering. The two data types are modeled as the two sets of vertices of a weighted bipartite graph. We then propose Isoperimetric Co-clustering Algorithm (ICA) - a new method for partitioning the bipartite graph. ICA requires a simple solution to a sparse system of linear equations instead of the eigenvalue or SVD problem in the popular spectral coclustering approach. Our theoretical analysis and extensive experiments performed on publicly available datasets demonstrate the advantages of ICA over other approaches in terms of the quality, efficiency and stability in partitioning the bipartite graph.
Description: “The original publication is available at www.springerlink.com.
Record URI: http://hdl.handle.net/1850/8223
Date: 2008

Files in this item

Files Size Format View
MregeArticle06-2008.pdf 788.6Kb PDF View/Open

The following license files are associated with this item:

This item appears in the following Collection(s)

Show full item record

Search RIT DML


Advanced Search

Browse