Enhancement of data clustering using TSS-DBSCAN approach for data mining
Cheng-Fa Tsai, Yao Chiang
Abstract
Cheng-Fa Tsai, Yao Chiang
Abstract
This work develops a new density-based clustering scheme, TSS-DBSCAN, which uses DBSCAN and a new method of applying two-phase screening, to reduce the extent of the meaningless expansion of clustering to improve data clustering for numerous related applications. Experimental results demonstrate that the proposed new TSS-DBSCAN scheme has very high noise filtering rate and clustering accuracy (both close to 100%), and is faster than some prominent density-based clustering methods, including KIDBSCAN, DBSCAN, IDBSCAN, QIDBSCAN, and SPY_DBSCAN. The presented approach may be the best density-based clustering method with low time cost in the world currently.
OpenAlex reports 4 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
This work develops a new density-based clustering scheme, TSS-DBSCAN, which uses DBSCAN and a new method of applying two-phase screening, to reduce the extent of the meaningless expansion of clustering to improve data clustering for numerous related applications. Experimental results demonstrate that the proposed new TSS-DBSCAN scheme has very high noise filtering rate and clustering accuracy (both close to 100%), and is faster than some prominent density-based clustering methods, including KIDBSCAN, DBSCAN, IDBSCAN, QIDBSCAN, and SPY_DBSCAN. The presented approach may be the best density-based clustering method with low time cost in the world currently.
Key concepts: DBSCAN, Cluster analysis, Computer science, Noise (video), Data mining, Pattern recognition (psychology), Scheme (mathematics), CURE data clustering algorithm