Fairness by Explicability and Adversarial SHAP Learning

The ability to understand and trust the fairness of model predictions, particularly when considering the outcomes of unprivileged groups, is critical to the deployment and adoption of machine learning systems. SHAP values provide a unified framework for interpreting model predictions and feature attribution but do not address the problem of fairness directly. In this work, we propose a new defi…

Paper

Similar papers

© 2026 NYSGPT2525 LLC