SPDFusion: An Infrared and Visible Image Fusion Network Based on a Non-Euclidean Representation of Riemannian Manifolds

Kang, Huan; Li, Hui; Xu, Tianyang; Wang, Rui; Wu, Xiao-Jun; Kittler, Josef

Computer Science > Computer Vision and Pattern Recognition

arXiv:2411.10679 (cs)

[Submitted on 16 Nov 2024 (v1), last revised 9 Mar 2025 (this version, v2)]

Title:SPDFusion: An Infrared and Visible Image Fusion Network Based on a Non-Euclidean Representation of Riemannian Manifolds

Authors:Huan Kang, Hui Li, Tianyang Xu, Rui Wang, Xiao-Jun Wu, Josef Kittler

View PDF HTML (experimental)

Abstract:Euclidean representation learning methods have achieved commendable results in image fusion tasks, which can be attributed to their clear advantages in handling with linear space. However, data collected from a realistic scene usually have a non-Euclidean structure, where Euclidean metric might be limited in representing the true data relationships, degrading fusion performance. To address this issue, a novel SPD (symmetric positive definite) manifold learning framework is proposed for multi-modal image fusion, named SPDFusion, which extends the image fusion approach from the Euclidean space to the SPD manifolds. Specifically, we encode images according to the Riemannian geometry to exploit their intrinsic statistical correlations, thereby aligning with human visual perception. Actually, the SPD matrix underpins our network learning, with a cross-modal fusion strategy employed to harness modality-specific dependencies and augment complementary information. Subsequently, an attention module is designed to process the learned weight matrix, facilitating the weighting of spatial global correlation semantics via SPD matrix multiplication. Based on this, we design an end-to-end fusion network based on cross-modal manifold learning. Extensive experiments on public datasets demonstrate that our framework exhibits superior performance compared to the current state-of-the-art methods.

Comments:	14 pages, 12 figures
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
ACM classes:	I.4
Cite as:	arXiv:2411.10679 [cs.CV]
	(or arXiv:2411.10679v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2411.10679

Submission history

From: Huan Kang [view email]
[v1] Sat, 16 Nov 2024 03:09:49 UTC (4,127 KB)
[v2] Sun, 9 Mar 2025 15:12:15 UTC (4,127 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:SPDFusion: An Infrared and Visible Image Fusion Network Based on a Non-Euclidean Representation of Riemannian Manifolds

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:SPDFusion: An Infrared and Visible Image Fusion Network Based on a Non-Euclidean Representation of Riemannian Manifolds

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators