To Which Out-Of-Distribution Object Orientations Are DNNs Capable of Generalizing?

Cooper, Avi; Boix, Xavier; Harari, Daniel; Madan, Spandan; Pfister, Hanspeter; Sasaki, Tomotake; Sinha, Pawan

Computer Science > Computer Vision and Pattern Recognition

arXiv:2109.13445v1 (cs)

[Submitted on 28 Sep 2021 (this version), latest version 13 Jul 2023 (v2)]

Title:To Which Out-Of-Distribution Object Orientations Are DNNs Capable of Generalizing?

Authors:Avi Cooper, Xavier Boix, Daniel Harari, Spandan Madan, Hanspeter Pfister, Tomotake Sasaki, Pawan Sinha

View PDF

Abstract:The capability of Deep Neural Networks (DNNs) to recognize objects in orientations outside the distribution of the training data, ie. out-of-distribution (OoD) orientations, is not well understood. For humans, behavioral studies showed that recognition accuracy varies across OoD orientations, where generalization is much better for some orientations than for others. In contrast, for DNNs, it remains unknown how generalization abilities are distributed among OoD orientations. In this paper, we investigate the limitations of DNNs' generalization capacities by systematically inspecting patterns of success and failure of DNNs across OoD orientations. We use an intuitive and controlled, yet challenging learning paradigm, in which some instances of an object category are seen at only a few geometrically restricted orientations, while other instances are seen at all orientations. The effect of data diversity is also investigated by increasing the number of instances seen at all orientations in the training set. We present a comprehensive analysis of DNNs' generalization abilities and limitations for representative architectures (ResNet, Inception, DenseNet and CORnet). Our results reveal an intriguing pattern -- DNNs are only capable of generalizing to instances of objects that appear like 2D, ie. in-plane, rotations of in-distribution orientations.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG); Neurons and Cognition (q-bio.NC); Machine Learning (stat.ML)
Cite as:	arXiv:2109.13445 [cs.CV]
	(or arXiv:2109.13445v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2109.13445

Submission history

From: Xavier Boix [view email]
[v1] Tue, 28 Sep 2021 02:48:00 UTC (16,065 KB)
[v2] Thu, 13 Jul 2023 04:23:23 UTC (44,959 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:To Which Out-Of-Distribution Object Orientations Are DNNs Capable of Generalizing?

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:To Which Out-Of-Distribution Object Orientations Are DNNs Capable of Generalizing?

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators