arXiv · 2506.06243
Fairmetrics: An R package for group fairness evaluation
Abstract
Fairness is a growing area of machine learning (ML) that focuses on ensuring models do not produce systematically biased outcomes for specific groups, particularly those defined by protected attributes such as race, gender, or age. Evaluating fairness is a critical aspect of ML model development, as biased models can perpetuate structural inequalities. The {fairmetrics} R package offers a user-friendly framework for rigorously evaluating numerous group-based fairness criteria, including metrics based on independence (e.g., statistical parity), separation (e.g., equalized odds), and sufficiency (e.g., predictive parity). Group-based fairness criteria assess whether a model is equally accurate or well-calibrated across a set of predefined groups so that appropriate bias mitigation strategies can be implemented. {fairmetrics} provides both point and interval estimates for multiple metrics through a convenient wrapper function and includes an example dataset derived from the Medical Information Mart for Intensive Care, version II (MIMIC-II) database (Goldberger et al., 2000; Raffa, 2016).
Explore related subjects
Keep this discovery
Benjamin Smith, Jianhui Gao, Jessica Gronsbell. 2025-06-06. Fairmetrics: An R package for group fairness evaluation. https://doi.org/10.21105/joss.08497
Cite the original work for its findings. Save a collection to share your selection of sources.