← 返回论文检索
ICML 2025PosterAccept (poster)

Fairness Overfitting in Machine Learning: An Information-Theoretic Perspective

Firas Laakom, Haobo Chen, Jürgen Schmidhuber, Yuheng Bu

KAUST · University of Florida · KAUST GenAI Center; Swiss AI Lab · University of California, Santa Barbara

PDF 由论文原始站点提供,PaperCompass 不保存论文文件。

摘要

Despite substantial progress in promoting fairness in high-stake applications using machine learning models, existing methods often modify the training process, such as through regularizers or other interventions, but lack formal guarantees that fairness achieved during training will generalize to unseen data. Although overfitting with respect to prediction performance has been extensively studied, overfitting in terms of fairness loss has received far less attention. This paper proposes a theoretical framework for analyzing fairness generalization error through an information-theoretic lens. Our novel bounding technique is based on Efron–Stein inequality, which allows us to derive tight information-theoretic fairness generalization bounds with both Mutual Information (MI) and Conditional Mutual Information (CMI). Our empirical results validate the tightness and practical relevance of these bounds across diverse fairness-aware learning algorithms.Our framework offers valuable insights to guide the design of algorithms improving fairness generalization.