CausaLM: Causal Model Explanation Through Counterfactual Language Models

Last update: Jul 10, 2022

Overview

CausaLM: Causal Model Explanation Through Counterfactual Language Models

Authors:

Amir Feder, Nadav Oved, Uri Shalit, Roi Reichart

Abstract:

Understanding predictions made by deep neural networks is notoriously difficult, but also crucial to their dissemination. As all ML-based methods, they are as good as their training data, and can also capture unwanted biases. While there are tools that can help understand whether such biases exist, they do not distinguish between correlation and causation, and might be ill-suited for text-based models and for reasoning about high level language concepts. A key problem of estimating the causal effect of a concept of interest on a given model is that this estimation requires the generation of counterfactual examples, which is challenging with existing generation technology. To bridge that gap, we propose CausaLM, a framework for producing causal model explanations using counterfactual language representation models. Our approach is based on fine-tuning of deep contextualized embedding models with auxiliary adversarial tasks derived from the causal graph of the problem. Concretely, we show that by carefully choosing auxiliary adversarial pre-training tasks, language representation models such as BERT can effectively learn a counterfactual representation for a given concept of interest, and be used to estimate its true causal effect on model performance. A byproduct of our method is a representation that is unaffected by the tested concept, which can be useful in mitigating unwanted bias ingrained in the data.

CausaLM: Causal Model Explanation Through Counterfactual Language Models

Related tags

Overview

CausaLM: Causal Model Explanation Through Counterfactual Language Models

Authors:

Amir Feder, Nadav Oved, Uri Shalit, Roi Reichart

Abstract:

Links:

Paper

Code

Data

Owner

Amir Feder

Implementation of Gans

Position detection system of mobile robot in the warehouse enviroment

ISTR: End-to-End Instance Segmentation with Transformers (https://arxiv.org/abs/2105.00637)

Face Mask Detector by live camera using tensorflow-keras, openCV and Python

This is a code repository for the paper "Graph Auto-Encoders for Financial Clustering".

Retinal Vessel Segmentation with Pixel-wise Adaptive Filters (ISBI 2022)

Image-generation-baseline - MUGE Text To Image Generation Baseline

Tensorflow2.0 🍎🍊 is delicious, just eat it! 😋😋

OpenGAN: Open-Set Recognition via Open Data Generation

Tooling for the Common Objects In 3D dataset.

Model Serving Made Easy

68 keypoint annotations for COFW test data

This repository contains various models targetting multimodal representation learning, multimodal fusion for downstream tasks such as multimodal sentiment analysis.

The CLRS Algorithmic Reasoning Benchmark

9th place solution

A knowledge base construction engine for richly formatted data

This repository contains the code used in the paper "Prompt-Based Multi-Modal Image Segmentation".

OpenLT: An open-source project for long-tail classification

Learning trajectory representations using self-supervision and programmatic supervision.

A rule learning algorithm for the deduction of syndrome definitions from time series data.