ODD: Overlap-aware Estimation of Model Performance under Distribution Shift

Mishra, Aayush; Liu, Anqi

Computer Science > Machine Learning

arXiv:2506.14978 (cs)

[Submitted on 17 Jun 2025]

Title:ODD: Overlap-aware Estimation of Model Performance under Distribution Shift

Authors:Aayush Mishra, Anqi Liu

View PDF HTML (experimental)

Abstract:Reliable and accurate estimation of the error of an ML model in unseen test domains is an important problem for safe intelligent systems. Prior work uses disagreement discrepancy (DIS^2) to derive practical error bounds under distribution shifts. It optimizes for a maximally disagreeing classifier on the target domain to bound the error of a given source classifier. Although this approach offers a reliable and competitively accurate estimate of the target error, we identify a problem in this approach which causes the disagreement discrepancy objective to compete in the overlapping region between source and target domains. With an intuitive assumption that the target disagreement should be no more than the source disagreement in the overlapping region due to high enough support, we devise Overlap-aware Disagreement Discrepancy (ODD). Maximizing ODD only requires disagreement in the non-overlapping target domain, removing the competition. Our ODD-based bound uses domain-classifiers to estimate domain-overlap and better predicts target performance than DIS^2. We conduct experiments on a wide array of benchmarks to show that our method improves the overall performance-estimation error while remaining valid and reliable. Our code and results are available on GitHub.

Comments:	Accepted to the 41st Conference on Uncertainty in Artificial Intelligence, 2025
Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:2506.14978 [cs.LG]
	(or arXiv:2506.14978v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2506.14978

Submission history

From: Aayush Mishra [view email]
[v1] Tue, 17 Jun 2025 21:05:42 UTC (796 KB)

Computer Science > Machine Learning

Title:ODD: Overlap-aware Estimation of Model Performance under Distribution Shift

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:ODD: Overlap-aware Estimation of Model Performance under Distribution Shift

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators