S-Divergence-Based Internal Clustering Validation Index

Kumar Sharma, Krishna; Seal, Ayan; Yazidi, Anis; Krejcar, Ondrej

Autor:

Kumar Sharma, Krishna

;

Seal, Ayan

;

Yazidi, Anis

;

Krejcar, Ondrej

Fecha:

12/2023

Palabra clave:

clustering quality indexes; generalized mean; K-Nearest Neighbors; S-distance; S-divergence; spectral clustering; symmetry favored; IJIMAI

Revista / editorial:

International Journal of Interactive Multimedia and Artificial Intelligence

Citación:

K. K. Sharma, A. Seal, A. Yazidi, O. Krejcar. S-Divergence-Based Internal Clustering Validation Index, International Journal of Interactive Multimedia and Artificial Intelligence, (2023), http://dx.doi.org/10.9781/ijimai.2023.10.001

Tipo de Ítem:

article

Resumen:

A clustering validation index (CVI) is employed to evaluate an algorithm’s clustering results. Generally, CVI statistics can be split into three classes, namely internal, external, and relative cluster validations. Most of the existing internal CVIs were designed based on compactness (CM) and separation (SM). The distance between cluster centers is calculated by SM, whereas the CM measures the variance of the cluster. However, the SM between groups is not always captured accurately in highly overlapping classes. In this article, we devise a novel internal CVI that can be regarded as a complementary measure to the landscape of available internal CVIs. Initially, a database’s clusters are modeled as a non-parametric density function estimated using kernel density estimation. Then the S-divergence (SD) and S-distance are introduced for measuring the SM and the CM, respectively. The SD is defined based on the concept of Hermitian positive definite matrices applied to density functions. The proposed internal CVI (PM) is the ratio of CM to SM. The PM outperforms the legacy measures presented in the literature on both superficial and realistic databases in various scenarios, according to empirical results from four popular clustering algorithms, including fuzzy k-means, spectral clustering, density peak clustering, and density-based spatial clustering applied to noisy data.

Mostrar el registro completo del ítem

Ficheros en el ítem

Nombre: ijimai8_4_12.pdf

Tamaño: 4.059Mb

Formato: application/pdf

Ver/Abrir

Este ítem aparece en la(s) siguiente(s) colección(es)

vol. 8, nº 4, december 2023

Año
2012
2013
2014
2015
2016
2017
2018
2019
2020
2021
2022
2023
2024
2025

Vistas
0
0
0
0
0
0
0
0
0
0
0
41
124
36
201

Descargas
0
0
0
0
0
0
0
0
0
0
0
30
41
17
88