DOAJ Open Access 2023

Cluster Validity Index for Uncertain Data Based on a Probabilistic Distance Measure in Feature Space

Changwan Ko Jaeseung Baek Behnam Tavakkol Young-Seon Jeong

Abstrak

Cluster validity indices (CVIs) for evaluating the result of the optimal number of clusters are critical measures in clustering problems. Most CVIs are designed for typical data-type objects called certain data objects. Certain data objects only have a singular value and include no uncertainty, so they are assumed to be information-abundant in the real world. In this study, new CVIs for uncertain data, based on kernel probabilistic distance measures to calculate the distance between two distributions in feature space, are proposed for uncertain clusters with arbitrary shapes, sub-clusters, and noise in objects. By transforming original uncertain data into kernel spaces, the proposed CVI accurately measures the compactness and separability of a cluster for arbitrary cluster shapes and is robust to noise and outliers in a cluster. The proposed CVI was evaluated for diverse types of simulated and real-life uncertain objects, confirming that the proposed validity indexes in feature space outperform the pre-existing ones in the original space.

Topik & Kata Kunci

Penulis (4)

C

Changwan Ko

J

Jaeseung Baek

B

Behnam Tavakkol

Y

Young-Seon Jeong

Format Sitasi

Ko, C., Baek, J., Tavakkol, B., Jeong, Y. (2023). Cluster Validity Index for Uncertain Data Based on a Probabilistic Distance Measure in Feature Space. https://doi.org/10.3390/s23073708

Akses Cepat

PDF tidak tersedia langsung

Cek di sumber asli →
Lihat di Sumber doi.org/10.3390/s23073708
Informasi Jurnal
Tahun Terbit
2023
Sumber Database
DOAJ
DOI
10.3390/s23073708
Akses
Open Access ✓