Benchmarking of Clustering Validity Measures Revisited
C. J. Simpson et al.
What the paper says
Validation plays a crucial role in the clustering process. Many different internal validity indices exist for the purpose of determining the best clustering solution(s) from a given collection of candidates, for example, as produced by different algorithms or different algorithm hyper‐parameters. In this study, we present a comprehensive benchmark study of 26 internal validity indices, which includes highly popular classic indices as well as more recently developed ones. We adopted an enhanced revision of the methodology presented in Vendramin et al. (2010), developed here to address several shortcomings of this previous work. This overall new approach consists of three complementary custom‐tailored evaluation sub‐methodologies, each of which has been designed to assess specific aspects of an index's behavior while preventing potential biases of the other sub‐methodologies. Each sub‐methodology features two complementary measures of performance, alongside mechanisms that allow for an in‐depth investigation of more complex behaviors of the internal validity indices under study. Additionally, a new collection of 16,177 datasets has been produced, paired with eight widely used clustering algorithms, for a wider applicability scope and representation of more diverse clustering scenarios.
Evidence weight
Balanced mode · F 0.40 / M 0.15 / V 0.05 / R 0.40
| F · citation impact | 0.50 × 0.4 = 0.20 |
| M · momentum | 0.50 × 0.15 = 0.07 |
| V · venue signal | 0.50 × 0.05 = 0.03 |
| R · text relevance † | 0.50 × 0.4 = 0.20 |
† Text relevance is estimated at 0.50 on the detail page — for your query’s actual relevance score, open this paper from a search result.