Skip to content
Open access

Density-Peak-Based Clustering in Reduced Feature Spaces

Aug 2026 · Dhaka University Journal of Science · 0 citations · 37 references

TL;DR

This work proposed to use Regularised Multidimensional Scaling using Radial Basis Function (RBF-MDS) for dimension reduction, a multidimensional scaling that mitigates the impact of irrelevant or redundant features, enabling 2D/3D visualisation of clusters for interpretability.

Abstract

Clustering is a fundamental data mining technique that groups data points by similarity. A critical challenge for clustering algorithms is the effective selection of initial cluster centers, often done through inefficient trial-and-error. To address this, a novel Adaptive Cluster Center Initialization using Density Peak for Geodesic Distance-based Clustering (AGDPC) method intelligently identifies optimal centers. It builds upon the Density Peaks Clustering (DPC) and Geodesic-Based Initialization (GDPC) approaches, using weighted Euclidean distances that incorporate the Pearson correlation coefficient. Furthermore, AGDPC adaptively optimizes its threshold parameter using data field density estimation entropy, enhancing its accuracy and automation. It should be noted that, AGDPC currently operates directly on high-dimensional data without dimensionality reduction. While this preserves the original structure of the data, high-dimensional spaces often suffer from the “curse of dimensionality,” where distances between points become less meaningful, and computational complexity increases. Introducing Multidimensional Scaling (MDS) as a pre-processing step can address these challenges by projecting data into a lower-dimensional space while preserving pairwise distances or dissimilarities as much as possible. In this work, we propose to use multidimensional scaling that mitigates the impact of irrelevant or redundant features, enabling 2D/3D visualisation of clusters for interpretability. We proposed to use Regularised Multidimensional Scaling using Radial Basis Function (RBF-MDS) for dimension reduction. The idea is to decrease its dimensionality, followed by the application of the adaptive strategy on the dataset in its reduced dimension. Our approaches are tested on various benchmarking datasets, and the outcomes are contrasted with those of DPC, GDPC, and one of the most conventional clustering methods, K-Means clustering. When tested on actual data, the experimental findings show that the proposed methods surpassed current leading approaches based on several clustering validation metrics and reduced computational time significantly. Dhaka Univ. J. Sci. 74(2): 271–282, 2026 (July)

Read PDF

Similar papers

Preprint Oct 2026

Gradient-Guided Density Peak Clustering

Density peak clustering (DPC) connects each observation to its nearest neighbor of higher density and identifies cluster centers as high-density observations with unusually large nearest neighbor uphill shifts. The resulting uphill paths from observations to cluster centers, however, can be irregular and unstable in lo...

Yikun Zhang, Yen-Chi Chen · 0 citations
Preprint Aug 2026

Robust K-means Clustering using the Density Power Divergence Measure

A robust clustering method that estimates cluster centers and covariance matrices using density power divergence measures combined with Mahalanobis distance, making it resistant to outliers and adaptable to heterogeneous, elliptical clusters, unlike the classical K-means algorithm is introduced.

Anirban Mondal, Paromita Banerjee, A. Mandal · 0 citations
Conference Aug 2026

Optimal Cluster Selection in Unsupervised Machine Learning Using K-Means Clustering

In unsupervised learning, Clustering is a core method used to determine unseen arrangements and structures within datasets by grouping similar instances together. Among the many clustering algorithms, Among clustering techniques, K-Means continues to be one of the most popular owing to its ease of implementation, fast...

I. Khan, H. Daud, Rajalingam Sokkalingam et al. · 0 citations
Review Open access Sep 2026

Clustering Spectral Line Cubes: A Multiview Prototype Approach

We present a new clustering paradigm to segment radio astronomy cubes based on unsupervised machine learning provided by a recent modification of self-organizing maps (SOMs) known as self-organizing uniform manifold approximation and projection (SOUMAP). SOUMAP simultaneously reduces sample size via neural vector quant...

Josh Taylor, S. Offner, R. Friesen et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.