Leveraging Labeled and Unlabeled Data for Consistent Fair Binary Classification

Abstract : We study the problem of fair binary classification using the notion of Equal Opportunity. It requires the true positive rate to distribute equally across the sensitive groups. Within this setting we show that the fair optimal classifier is obtained by recalibrating the Bayes classifier by a group-dependent threshold. We provide a constructive expression for the threshold. This result motivates us to devise a plug-in classification procedure based on both unlabeled and labeled datasets. While the latter is used to learn the output conditional probability, the former is used for calibration. The overall procedure can be computed in polynomial time and it is shown to be statistically consistent both in terms of classification error and fairness measure. Finally, we present numerical experiments which indicate that our method is often superior or competitive with the state-of-the-art methods on benchmark datasets.
Complete list of metadatas

https://hal-upec-upem.archives-ouvertes.fr/hal-02150662
Contributor : Mohamed Hebiri <>
Submitted on : Friday, June 7, 2019 - 12:59:48 PM
Last modification on : Friday, October 4, 2019 - 1:25:50 AM

Files

main.pdf
Files produced by the author(s)

Identifiers

  • HAL Id : hal-02150662, version 1
  • ARXIV : 1906.05082

Collections

Citation

Evgenii Chzhen, Christophe Denis, Mohamed Hebiri, Luca Oneto, Massimiliano Pontil. Leveraging Labeled and Unlabeled Data for Consistent Fair Binary Classification. 2019. ⟨hal-02150662⟩

Share

Metrics

Record views

60

Files downloads

68