Characterizing Out-of-Distribution Error via Optimal Transport

Lu, Yuzhe; Qin, Yilong; Zhai, Runtian; Shen, Andrew; Chen, Ketong; Wang, Zhenlin; Kolouri, Soheil; Stepputtis, Simon; Campbell, Joseph; Sycara, Katia

Computer Science > Machine Learning

arXiv:2305.15640 (cs)

[Submitted on 25 May 2023 (v1), last revised 27 Oct 2023 (this version, v3)]

Title:Characterizing Out-of-Distribution Error via Optimal Transport

Authors:Yuzhe Lu, Yilong Qin, Runtian Zhai, Andrew Shen, Ketong Chen, Zhenlin Wang, Soheil Kolouri, Simon Stepputtis, Joseph Campbell, Katia Sycara

View PDF

Abstract:Out-of-distribution (OOD) data poses serious challenges in deployed machine learning models, so methods of predicting a model's performance on OOD data without labels are important for machine learning safety. While a number of methods have been proposed by prior work, they often underestimate the actual error, sometimes by a large margin, which greatly impacts their applicability to real tasks. In this work, we identify pseudo-label shift, or the difference between the predicted and true OOD label distributions, as a key indicator to this underestimation. Based on this observation, we introduce a novel method for estimating model performance by leveraging optimal transport theory, Confidence Optimal Transport (COT), and show that it provably provides more robust error estimates in the presence of pseudo-label shift. Additionally, we introduce an empirically-motivated variant of COT, Confidence Optimal Transport with Thresholding (COTT), which applies thresholding to the individual transport costs and further improves the accuracy of COT's error estimates. We evaluate COT and COTT on a variety of standard benchmarks that induce various types of distribution shift -- synthetic, novel subpopulation, and natural -- and show that our approaches significantly outperform existing state-of-the-art methods with an up to 3x lower prediction error.

Comments:	NeurIPS 2023
Subjects:	Machine Learning (cs.LG); Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2305.15640 [cs.LG]
	(or arXiv:2305.15640v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2305.15640

Submission history

From: Yuzhe Lu [view email]
[v1] Thu, 25 May 2023 01:37:13 UTC (3,711 KB)
[v2] Sat, 27 May 2023 01:08:15 UTC (1,857 KB)
[v3] Fri, 27 Oct 2023 21:33:27 UTC (1,861 KB)

Computer Science > Machine Learning

Title:Characterizing Out-of-Distribution Error via Optimal Transport

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Characterizing Out-of-Distribution Error via Optimal Transport

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators