Sciweavers

Free Online Productivity Tools i2Speak i2Symbol i2OCR iTex2Img iWeb2Print iWeb2Shot i2Type iPdf2Split iPdf2Merge i2Bopomofo i2Arabic i2Style i2Image i2PDF iLatex2Rtf Sci2ools

196

ICDCS
2006
IEEE

258views Distributed And Parallel Com...» more ICDCS 2006»

ParRescue: Scalable Parallel Algorithm and Implementation for Biclustering over Large Distributed Datasets

16 years 18 days ago

ParRescue: Scalable Parallel Algorithm and Implementation for Biclustering over Large Distributed Datasets

Download multimedia.ece.uic.edu

Biclustering refers to simultaneously capturing correlations present among subsets of attributes (columns) and records (rows). It is widely used in data mining applications including biological data analysis, ﬁnancial forecasting, and text mining. Biclustering algorithms are signiﬁcantly more complex compared to the classical one dimensional clustering techniques, particularly those requiring multiple computing platforms for large and distributed data sets. In this paper, we develop an efﬁcient scalable algorithm, referred to as ParRescue(Parallel Residue Co-clustering), that is capable of performing biclustering on extremely large or geographically distributed data sets. ParRescue divides the cluster tasks among processors with minimal communication costs thus making it scalable over large number of computing nodes. The proposed implementation is based on an existing sequential approach that has been modiﬁed for amenable parallel implementation. The proposed ParRescue algorit...

Jianhong Zhou, Ashfaq A. Khokhar

Real-time Traffic

Data Mining Applications | Distributed And Parallel Computing | Distributed Data Sets | Efﬁcient Scalable Algorithm | ICDCS 2006 |

claim paper

Related Content

» High Performance ParallelDistributed Biclustering Using Barycenter Heuristic

» A Scalable Parallel Approach for Peptide Identification from LargeScale Mass Spectrometry ...

» A Distributed Kernel Summation Framework for GeneralDimension Machine Learning

» Design and analysis of a multidimensional data sampling service for large scale data analy...

» MOVE A Large Scale KeywordBased Content Filtering and Dissemination System

» Scalable Algorithms for Distribution Search

» A Gaussian Belief Propagation Solver for Large Scale Support Vector Machines

» Towards billionbit optimization via a parallel estimation of distribution algorithm

» A Scalable Parallel Subspace Clustering Algorithm for Massive Data Sets

Post Info
More Details (n/a)

Added	11 Jun 2010
Updated	11 Jun 2010
Type	Conference
Year	2006
Where	ICDCS
Authors	Jianhong Zhou, Ashfaq A. Khokhar

Comments (0)