Masked Generative Extractor for Synergistic Representation and 3D Generation of Point Clouds

Zeng, Hongliang; Zhang, Ping; Li, Fang; Wang, Jiahua; Ye, Tingyu; Guo, Pengteng

Computer Science > Computer Vision and Pattern Recognition

arXiv:2406.17342 (cs)

[Submitted on 25 Jun 2024 (v1), last revised 15 Aug 2024 (this version, v2)]

Title:Masked Generative Extractor for Synergistic Representation and 3D Generation of Point Clouds

Authors:Hongliang Zeng, Ping Zhang, Fang Li, Jiahua Wang, Tingyu Ye, Pengteng Guo

View PDF HTML (experimental)

Abstract:Representation and generative learning, as reconstruction-based methods, have demonstrated their potential for mutual reinforcement across various domains. In the field of point cloud processing, although existing studies have adopted training strategies from generative models to enhance representational capabilities, these methods are limited by their inability to genuinely generate 3D shapes. To explore the benefits of deeply integrating 3D representation learning and generative learning, we propose an innovative framework called \textit{Point-MGE}. Specifically, this framework first utilizes a vector quantized variational autoencoder to reconstruct a neural field representation of 3D shapes, thereby learning discrete semantic features of point patches. Subsequently, we design a sliding masking ratios to smooth the transition from representation learning to generative learning. Moreover, our method demonstrates strong generalization capability in learning high-capacity models, achieving new state-of-the-art performance across multiple downstream tasks. In shape classification, Point-MGE achieved an accuracy of 94.2% (+1.0%) on the ModelNet40 dataset and 92.9% (+5.5%) on the ScanObjectNN dataset. Experimental results also confirmed that Point-MGE can generate high-quality 3D shapes in both unconditional and conditional settings.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2406.17342 [cs.CV]
	(or arXiv:2406.17342v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2406.17342

Submission history

From: Hongliang Zeng [view email]
[v1] Tue, 25 Jun 2024 07:57:03 UTC (14,313 KB)
[v2] Thu, 15 Aug 2024 09:59:58 UTC (20,923 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Masked Generative Extractor for Synergistic Representation and 3D Generation of Point Clouds

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Masked Generative Extractor for Synergistic Representation and 3D Generation of Point Clouds

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators