Scalable Techniques for Clustering the Web.
Data up to Jan 2025
Total Citations Per Year
Abstract
References (15)
Introduction to Modern Information Retrieval
1983 • 9,967 citations
An algorithm for suffix stripping
1980 • 8,166 citations
Approximate nearest neighbors
1998 • 4,060 citations
Similarity Search in High Dimensions via Hashing
1999 • 3,202 citations
On the resemblance and containment of documents
2002 • 1,695 citations
Syntactic clustering of the Web
1997 • 1,378 citations
Web document clustering
1998 • 1,093 citations
A Best Possible Heuristic for the k-Center Problem
1985 • 918 citations
Min-Wise Independent Permutations
2000 • 848 citations
Learning to extract symbolic knowledge from the World Wide Web
1998 • 679 citations
Computing Iceberg Queries Efficiently
1998 • 359 citations
Finding interesting associations without support pruning
2001 • 354 citations
WebBase: a repository of Web pages
2000 • 194 citations
A Small Approximately Min-Wise Independent Family of Hash Functions
2001 • 185 citations
Detecting digital copyright violations on the internet
1999 • 27 citations
Cited By (0)
No citing papers found in database