Sciweavers

Free Online Productivity Tools i2Speak i2Symbol i2OCR iTex2Img iWeb2Print iWeb2Shot i2Type iPdf2Split iPdf2Merge i2Bopomofo i2Arabic i2Style i2Image i2PDF iLatex2Rtf Sci2ools

150

Voted

ECAI
2010
Springer

197views Artificial Intelligence» more ECAI 2010»

A Very Fast Method for Clustering Big Text Datasets

15 years 6 months ago

A Very Fast Method for Clustering Big Text Datasets

Download www.cs.cmu.edu

Large-scale text datasets have long eluded a family of particularly elegant and effective clustering methods that exploits the power of pair-wise similarities between data points due to the prohibitive cost, time- and space-wise, in operating on a similarity matrix, where the state-of-the-art is at best quadratic in time and in space. We present an extremely fast and simple method also using the power of all pair-wise similarity between data points, and show through experiments that it does as well as previous methods in clustering accuracy, and it does so with in linear time and space, without sampling data points or sparsifying the similarity matrix.

Frank Lin, William W. Cohen

Real-time Traffic

Artificial Intelligence | Data Points | ECAI 2010 | Effective Clustering Methods | Similarity Matrix |

claim paper

Related Content

» Hierarchical Density Shaving A clustering and visualization framework for large biological...

» CRD fast coclustering on large datasets utilizing samplingbased matrix decomposition

» Patch Relational Neural Gas Clustering of Huge Dissimilarity Datasets

» Evaluation of text clustering methods using wordnet

» An objective evaluation criterion for clustering

» CBC Clustering Based Text Classification Requiring Minimal Labeled Data

» Handwritten Arabic text line segmentation using affinity propagation

» Power Iteration Clustering

» CLOPE a fast and effective clustering algorithm for transactional data

Post Info
More Details (n/a)

Added	08 Nov 2010
Updated	08 Nov 2010
Type	Conference
Year	2010
Where	ECAI
Authors	Frank Lin, William W. Cohen

Comments (0)