K-means is a simple algorithm but it's not fast on GPUs
Flash-KMeans, an IO-aware implementation of exact k-means, redesigned to address the bottlenecks of modern GPUs.
By directly addressing memory bottlenecks:
- up to 30x faster than cuML
- up to 200x faster than FAISS
- while using the same algorithm, just optimized for modern hardware
On scales of millions of points, a single iteration of k-means takes just milliseconds.
A classic algorithm, redesigned for modern GPUs. Paper, Code
••••••••••••••••••••••••••••••••••••••
🤖 Data & ML | @DataXplore
Post #2091
191