This repository provides CLUBench, a comprehensive benchmark for tabular clustering.
Paper: [CLUBench: A Clustering Benchmark]
Code: CLUBench
Clustering is a fundamental problem in data science with a long-standing research history. Over the past decades, numerous clustering algorithms, ranging from conventional machine learning approaches to deep clustering methods, have been developed. Despite this progress, a systematic and… See the full description on the dataset page:
https://huggingface.co/datasets/Feng-001/Clustering-Benchmark.