Inferring Unfairness and Error from Population Statistics in Binary and Multiclass Classification

Sabato, Sivan; Treister, Eran; Yom-Tov, Elad

Computer Science > Machine Learning

arXiv:2206.03234v1 (cs)

[Submitted on 7 Jun 2022 (this version), latest version 5 Apr 2024 (v2)]

Title:Inferring Unfairness and Error from Population Statistics in Binary and Multiclass Classification

Authors:Sivan Sabato, Eran Treister, Elad Yom-Tov

View PDF

Abstract:We propose methods for making inferences on the fairness and accuracy of a given classifier, using only aggregate population statistics. This is necessary when it is impossible to obtain individual classification data, for instance when there is no access to the classifier or to a representative individual-level validation set. We study fairness with respect to the equalized odds criterion, which we generalize to multiclass classification. We propose a measure of unfairness with respect to this criterion, which quantifies the fraction of the population that is treated unfairly. We then show how inferences on the unfairness and error of a given classifier can be obtained using only aggregate label statistics such as the rate of prediction of each label in each sub-population, as well as the true rate of each label. We derive inference procedures for binary classifiers and for multiclass classifiers, for the case where confusion matrices in each sub-population are known, and for the significantly more challenging case where they are unknown. We report experiments on data sets representing diverse applications, which demonstrate the effectiveness and the wide range of possible uses of the proposed methodology.

Subjects:	Machine Learning (cs.LG); Computers and Society (cs.CY); Machine Learning (stat.ML)
Cite as:	arXiv:2206.03234 [cs.LG]
	(or arXiv:2206.03234v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2206.03234

Submission history

From: Sivan Sabato [view email]
[v1] Tue, 7 Jun 2022 12:26:28 UTC (923 KB)
[v2] Fri, 5 Apr 2024 18:00:01 UTC (670 KB)

Computer Science > Machine Learning

Title:Inferring Unfairness and Error from Population Statistics in Binary and Multiclass Classification

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Inferring Unfairness and Error from Population Statistics in Binary and Multiclass Classification

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators