Bridging the inference gap in Mutimodal Variational Autoencoders

Senellart, Agathe; Allassonnière, Stéphanie

Computer Science > Machine Learning

arXiv:2502.03952 (cs)

[Submitted on 6 Feb 2025]

Title:Bridging the inference gap in Mutimodal Variational Autoencoders

Authors:Agathe Senellart, Stéphanie Allassonnière

View PDF HTML (experimental)

Abstract:From medical diagnosis to autonomous vehicles, critical applications rely on the integration of multiple heterogeneous data modalities. Multimodal Variational Autoencoders offer versatile and scalable methods for generating unobserved modalities from observed ones. Recent models using mixturesof-experts aggregation suffer from theoretically grounded limitations that restrict their generation quality on complex datasets. In this article, we propose a novel interpretable model able to learn both joint and conditional distributions without introducing mixture aggregation. Our model follows a multistage training process: first modeling the joint distribution with variational inference and then modeling the conditional distributions with Normalizing Flows to better approximate true posteriors. Importantly, we also propose to extract and leverage the information shared between modalities to improve the conditional coherence of generated samples. Our method achieves state-of-the-art results on several benchmark datasets.

Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:2502.03952 [cs.LG]
	(or arXiv:2502.03952v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2502.03952

Submission history

From: Agathe Senellart [view email]
[v1] Thu, 6 Feb 2025 10:43:55 UTC (28,024 KB)

Computer Science > Machine Learning

Title:Bridging the inference gap in Mutimodal Variational Autoencoders

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Bridging the inference gap in Mutimodal Variational Autoencoders

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators