Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Counterfactual Models for Fair and Adequate Explanations

View through CrossRef
Recent efforts have uncovered various methods for providing explanations that can help interpret the behavior of machine learning programs. Exact explanations with a rigorous logical foundation provide valid and complete explanations, but they have an epistemological problem: they are often too complex for humans to understand and too expensive to compute even with automated reasoning methods. Interpretability requires good explanations that humans can grasp and can compute. We take an important step toward specifying what good explanations are by analyzing the epistemically accessible and pragmatic aspects of explanations. We characterize sufficiently good, or fair and adequate, explanations in terms of counterfactuals and what we call the conundra of the explainee, the agent that requested the explanation. We provide a correspondence between logical and mathematical formulations for counterfactuals to examine the partiality of counterfactual explanations that can hide biases; we define fair and adequate explanations in such a setting. We provide formal results about the algorithmic complexity of fair and adequate explanations. We then detail two sophisticated counterfactual models, one based on causal graphs, and one based on transport theories. We show transport based models have several theoretical advantages over the competition as explanation frameworks for machine learning algorithms.
Title: Counterfactual Models for Fair and Adequate Explanations
Description:
Recent efforts have uncovered various methods for providing explanations that can help interpret the behavior of machine learning programs.
Exact explanations with a rigorous logical foundation provide valid and complete explanations, but they have an epistemological problem: they are often too complex for humans to understand and too expensive to compute even with automated reasoning methods.
Interpretability requires good explanations that humans can grasp and can compute.
We take an important step toward specifying what good explanations are by analyzing the epistemically accessible and pragmatic aspects of explanations.
We characterize sufficiently good, or fair and adequate, explanations in terms of counterfactuals and what we call the conundra of the explainee, the agent that requested the explanation.
We provide a correspondence between logical and mathematical formulations for counterfactuals to examine the partiality of counterfactual explanations that can hide biases; we define fair and adequate explanations in such a setting.
We provide formal results about the algorithmic complexity of fair and adequate explanations.
We then detail two sophisticated counterfactual models, one based on causal graphs, and one based on transport theories.
We show transport based models have several theoretical advantages over the competition as explanation frameworks for machine learning algorithms.

Related Results

The Counterfactual Analysis in EU Merger Control
The Counterfactual Analysis in EU Merger Control
The counterfactual method, which can be used to assess the effects of an actual or a hypothetical event, has always played an important role in EU competition law. It has recently...
Explaining Data-Driven Decisions made by AI Systems: The Counterfactual Approach
Explaining Data-Driven Decisions made by AI Systems: The Counterfactual Approach
We examine counterfactual explanations for explaining the decisions made by model-based AI systems. The counterfactual approach we consider defines an explanation as a set of the s...
Counterfactual Examples for Data Augmentation: A Case Study
Counterfactual Examples for Data Augmentation: A Case Study
Counterfactual explanations are gaining in popularity as a way of explaining machine learning models. Counterfactual examples are generally created to help interpret the decision o...
Data Augmentation using Counterfactuals: Proximity vs Diversity
Data Augmentation using Counterfactuals: Proximity vs Diversity
Counterfactual explanations are gaining in popularity as a way of explaining machine learning models. Counterfactual examples are generally created to help interpret the decision o...
FAIR-IMPACT
FAIR-IMPACT
In this poster we present the FAIR-IMPACT project, “Expanding FAIR solutions across EOSC”, which is funded by the European Commission Horizon Europe programme. The acronym FAIR st...
A Capacity Building Program for developing FAIR skills
A Capacity Building Program for developing FAIR skills
GO FAIR is an international, bottom-up movement dedicated to adhering as closely as possible to the FAIR Guiding Principles in the implementation of data and services as outlined i...
Downward counterfactual insights into weather extremes
Downward counterfactual insights into weather extremes
<p>There are many regions where the duration of reliable scientific observations of key weather hazard variables, such as rainfall and wind speed, is of the order of ...

Back to Top