Revisiting Dimensionality Reduction Techniques for Visual Cluster Analysis: An Empirical Study

Xia, Jiazhi; Zhang, Yuchen; Song, Jie; Chen, Yang; Wang, Yunhai; Liu, Shixia

doi:10.1109/TVCG.2021.3114694

Computer Science > Human-Computer Interaction

arXiv:2110.02894 (cs)

[Submitted on 6 Oct 2021]

Title:Revisiting Dimensionality Reduction Techniques for Visual Cluster Analysis: An Empirical Study

Authors:Jiazhi Xia, Yuchen Zhang, Jie Song, Yang Chen, Yunhai Wang, Shixia Liu

View PDF

Abstract:Dimensionality Reduction (DR) techniques can generate 2D projections and enable visual exploration of cluster structures of high-dimensional datasets. However, different DR techniques would yield various patterns, which significantly affect the performance of visual cluster analysis tasks. We present the results of a user study that investigates the influence of different DR techniques on visual cluster analysis. Our study focuses on the most concerned property types, namely the linearity and locality, and evaluates twelve representative DR techniques that cover the concerned properties. Four controlled experiments were conducted to evaluate how the DR techniques facilitate the tasks of 1) cluster identification, 2) membership identification, 3) distance comparison, and 4) density comparison, respectively. We also evaluated users' subjective preference of the DR techniques regarding the quality of projected clusters. The results show that: 1) Non-linear and Local techniques are preferred in cluster identification and membership identification; 2) Linear techniques perform better than non-linear techniques in density comparison; 3) UMAP (Uniform Manifold Approximation and Projection) and t-SNE (t-Distributed Stochastic Neighbor Embedding) perform the best in cluster identification and membership identification; 4) NMF (Nonnegative Matrix Factorization) has competitive performance in distance comparison; 5) t-SNLE (t-Distributed Stochastic Neighbor Linear Embedding) has competitive performance in density comparison.

Comments:	IEEE VIS 2021, to appear in IEEE Transactions on Visualization & Computer Graphics
Subjects:	Human-Computer Interaction (cs.HC)
Cite as:	arXiv:2110.02894 [cs.HC]
	(or arXiv:2110.02894v1 [cs.HC] for this version)
	https://doi.org/10.48550/arXiv.2110.02894
Related DOI:	https://doi.org/10.1109/TVCG.2021.3114694

Submission history

From: Jie Song [view email]
[v1] Wed, 6 Oct 2021 16:19:39 UTC (1,189 KB)

Computer Science > Human-Computer Interaction

Title:Revisiting Dimensionality Reduction Techniques for Visual Cluster Analysis: An Empirical Study

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Human-Computer Interaction

Title:Revisiting Dimensionality Reduction Techniques for Visual Cluster Analysis: An Empirical Study

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators