ConiVAT: Cluster Tendency Assessment and Clustering with Partial Background Knowledge

Rathore, Punit; Bezdek, James C.; Santi, Paolo; Ratti, Carlo

Computer Science > Machine Learning

arXiv:2008.09570v1 (cs)

[Submitted on 21 Aug 2020 (this version), latest version 28 Sep 2020 (v2)]

Title:ConiVAT: Cluster Tendency Assessment and Clustering with Partial Background Knowledge

Authors:Punit Rathore, James C. Bezdek, Paolo Santi, Carlo Ratti

View PDF

Abstract:The VAT method is a visual technique for determining the potential cluster structure and the possible number of clusters in numerical data. Its improved version, iVAT, uses a path-based distance transform to improve the effectiveness of VAT for "tough" cases. Both VAT and iVAT have also been used in conjunction with a single-linkage(SL) hierarchical clustering algorithm. However, they are sensitive to noise and bridge points between clusters in the dataset, and consequently, the corresponding VAT/iVAT images are often in-conclusive for such cases. In this paper, we propose a constraint-based version of iVAT, which we call ConiVAT, that makes use of background knowledge in the form of constraints, to improve VAT/iVAT for challenging and complex datasets. ConiVAT uses the input constraints to learn the underlying similarity metric and builds a minimum transitive dissimilarity matrix, before applying VAT to it. We demonstrate ConiVAT approach to visual assessment and single linkage clustering on nine datasets to show that, it improves the quality of iVAT images for complex datasets, and it also overcomes the limitation of SL clustering with VAT/iVAT due to "noisy" bridges between clusters. Extensive experiment results on nine datasets suggest that ConiVAT outperforms the other three semi-supervised clustering algorithms in terms of improved clustering accuracy.

Comments:	Submitted to IEEE Transactions on Knowledge and Data Engineering
Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:2008.09570 [cs.LG]
	(or arXiv:2008.09570v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2008.09570

Submission history

From: Punit Rathore [view email]
[v1] Fri, 21 Aug 2020 16:30:31 UTC (18,252 KB)
[v2] Mon, 28 Sep 2020 17:21:09 UTC (18,253 KB)

Computer Science > Machine Learning

Title:ConiVAT: Cluster Tendency Assessment and Clustering with Partial Background Knowledge

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:ConiVAT: Cluster Tendency Assessment and Clustering with Partial Background Knowledge

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators