Back to all papers

Reliable fairness auditing with semi-supervised inference.

July 1, 2026pubmed logopapers

Authors

Gao J,Gronsbell J

Affiliations (1)

  • Department of Statistical Science, University of Toronto, Toronto, ON M7A 2S4, Canada.

Abstract

Machine learning (ML) models often exhibit bias that can exacerbate inequities in biomedical applications. Fairness auditing, the process of evaluating a model's performance across subpopulations, is critical for identifying and mitigating these biases. However, audits typically rely on large volumes of labeled data, which are costly and labor-intensive to obtain. To address this challenge, we introduce Infairness, a unified framework for auditing a wide range of fairness criteria using semi-supervised inference. Our approach combines a small labeled dataset with a large unlabeled dataset by imputing missing outcomes via regression with carefully selected nonlinear basis functions. Through extensive theoretical and empirical analyses, we show that our proposed estimator is (1) robust to specification of the ML or imputation model and (2) substantially more efficient than supervised estimation based solely on the labeled data. In two real-world fairness audits using electronic health record and medical imaging data, Infairness reduces variance by 40 - 60% compared to supervised estimation, underscoring its value for reliable fairness auditing with limited labeled data.

Topics

Machine LearningModels, StatisticalJournal Article

Ready to Sharpen Your Edge?

Subscribe to join 11k+ peers who rely on RadAISlice. Get the essential weekly briefing that empowers you to navigate the future of radiology.

We respect your privacy. Unsubscribe at any time.