Almost Right: Making First-layer Kernels Nearly Orthogonal Improves Model Generalization

Crum, Colton R.; Czajka, Adam

Computer Science > Computer Vision and Pattern Recognition

arXiv:2504.16362 (cs)

[Submitted on 23 Apr 2025]

Title:Almost Right: Making First-layer Kernels Nearly Orthogonal Improves Model Generalization

Authors:Colton R. Crum, Adam Czajka

View PDF HTML (experimental)

Abstract:An ongoing research challenge within several domains in computer vision is how to increase model generalization capabilities. Several attempts to improve model generalization performance are heavily inspired by human perceptual intelligence, which is remarkable in both its performance and efficiency to generalize to unknown samples. Many of these methods attempt to force portions of the network to be orthogonal, following some observation within neuroscience related to early vision processes. In this paper, we propose a loss component that regularizes the filtering kernels in the first convolutional layer of a network to make them nearly orthogonal. Deviating from previous works, we give the network flexibility in which pairs of kernels it makes orthogonal, allowing the network to navigate to a better solution space, imposing harsh penalties. Without architectural modifications, we report substantial gains in generalization performance using the proposed loss against previous works (including orthogonalization- and saliency-based regularization methods) across three different architectures (ResNet-50, DenseNet-121, ViT-b-16) and two difficult open-set recognition tasks: presentation attack detection in iris biometrics, and anomaly detection in chest X-ray images.

Comments:	8 pages, 1 figure, 3 tables
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2504.16362 [cs.CV]
	(or arXiv:2504.16362v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2504.16362

Submission history

From: Colton Crum [view email]
[v1] Wed, 23 Apr 2025 02:27:20 UTC (523 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Almost Right: Making First-layer Kernels Nearly Orthogonal Improves Model Generalization

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Almost Right: Making First-layer Kernels Nearly Orthogonal Improves Model Generalization

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators