International Conference on Machine Learning · 2016 · 31 citations · 23 references
Mathematical ProgrammingCluster ComputingEngineeringComputational ComplexityRange SearchingUnsupervised Machine LearningData ScienceData MiningPattern RecognitionYinyang K-meansHolder InequalityCombinatorial OptimizationComputational GeometryApproximation TheoryLow-rank ApproximationManifold LearningComputer ScienceDimensionality ReductionVoronoi DiagramNonlinear Dimensionality ReductionGeometric AlgorithmBlock VectorsSimilarity Search
This paper introduces a new method to approximate Euclidean distances between points using block vectors in combination with the Holder inequality. By defining lower bounds based on the proposed approximation, cluster algorithms can be considerably accelerated without loss of quality. In extensive experiments, we show a considerable reduction in terms of computational time in comparison to standard methods and the recently proposed Yinyang k-means. Additionally we show that the memory consumption of the presented clustering algorithm does not depend on the number of clusters, which makes the approach suitable for large scale problems.
23
Least squares quantization in PCM
Sheelagh Lloyd · IEEE Transactions on Information Theory · 1982 · 15.1K citations · Full text
Top 10 algorithms in data mining
Xindong Wu, Vipin Kumar, J. R. Quinlan et al. · Knowledge and Information Systems · 2007 · 5.6K citations
Object retrieval with large vocabularies and fast spatial matching
James Philbin, Ondřej Chum, Michael Isard et al. · 2007 · 3K citations