Towards Debiasing NLU Models from Unknown Biases

Last update: Jun 14, 2022

Overview

Towards Debiasing NLU Models from Unknown Biases

Abstract: NLU models often exploit biased features to achieve high dataset-specific performance without properly learning the intended task. Recently proposed debiasing methods are shown to be effective in mitigating this tendency. However, these methods rely on a major assumption that the type of biased features is known a-priori, which limits their application to many NLU tasks and datasets. In this work, we present the first step to bridge this gap by introducing a self-debiasing framework that prevents models from mainly utilizing biases without knowing them in advance. The proposed framework is general and complementary to the existing debiasing methods. We show that the proposed framework allows these existing methods to retain the improvement on the challenge datasets (i.e., sets of examples designed to expose models’ reliance to biases) without specifically targeting certain biases. Furthermore, the evaluation suggests that applying the framework results in improved overall robustness.

The repository contains the code to reproduce our work in debiasing NLU models without prior information on biases. We provide 3 runs of experiment that are shown in our paper:

Debias MNLI model from syntactic bias and evaluate on HANS as the out-of-distribution data using example reweighting.
Debias MNLI model from syntactic bias and evaluate on HANS as the out-of-distribution data using product of expert.
Debias MNLI model from syntactic bias and evaluate on HANS as the out-of-distribution data using confidence regularization.

Requirements

The code requires python >= 3.6 and pytorch >= 1.1.0.

Additional required dependencies can be found in requirements.txt. Install all requirements by running:

pip install -r requirements.txt

Data

Our experiments use MNLI dataset version provided by GLUE benchmark. Download the file from here, and unzip under the directory ./dataset The dataset directory should be structured as the following:

└── dataset 
    └── MNLI
        ├── train.tsv
        ├── dev_matched.tsv
        ├── dev_mismatched.tsv
        ├── dev_mismatched.tsv

Running the experiments

For each evaluation setting, use the --mode arguments to set the appropriate loss function. Choose the annealed version of the loss function for reproducing the annealed results.

To reproduce our result on MNLI ⮕ HANS, run the following:

cd src/
CUDA_VISIBLE_DEVICES=9 python train_distill_bert.py \
  --output_dir ../experiments_self_debias_mnli_seed111/bert_reweighted_sampled2K_teacher_seed111_annealed_1to08 \
  --do_train --do_eval --mode reweight_by_teacher_annealed \
  --custom_teacher ../teacher_preds/mnli_trained_on_sample2K_seed111.json --seed 111 --which_bias hans

Biased examples identification

To obtain predictions of the shallow models, we train the same model architecture on the fraction of the dataset. For MNLI we subsample 2000 examples and train the model for 5 epochs. For obtaining shallow models of other datasets please see the appendix of our paper. The shallow model can be obtained with the command below:

cd src/
CUDA_VISIBLE_DEVICES=9 python train_distill_bert.py \
 --output_dir ../experiments_shallow_mnli/bert_base_sampled2K_seed111 \
 --do_train --do_eval --do_eval_on_train --mode none\
 --seed 111 --which_bias hans --debug --num_train_epochs 5 --debug_num 2000

Once the training and the evaluation on train set is done, copy the probability json files in the output directory to ../teacher_preds/mnli_trained_on_sample2K_seed111.json.

Expected results

Results on the MNLI ⮕ HANS setting without annealing:

Mode	Seed	MNLI-m	MNLI-mm	HANS avg.
None	111	84.57	84.72	62.04
reweighting	111	81.8	82.3	72.1
PoE	111	81.5	81.1	70.3
conf-reg	222	83.7	84.1	68.7

Towards Debiasing NLU Models from Unknown Biases

Related tags

Overview

Towards Debiasing NLU Models from Unknown Biases

Requirements

Data

Running the experiments

Biased examples identification

Expected results

Owner

Ubiquitous Knowledge Processing Lab

SpineAI Bilsky Grading With Python

50-days-of-Statistics-for-Data-Science - This repository consist of a 50-day program

GLM (General Language Model)

Repo for WWW 2022 paper: Progressively Optimized Bi-Granular Document Representation for Scalable Embedding Based Retrieval

Reference implementation for Deep Unsupervised Learning using Nonequilibrium Thermodynamics

Using Convolutional Neural Networks (CNN) for Semantic Segmentation of Breast Cancer Lesions (BRCA)

Radar-to-Lidar: Heterogeneous Place Recognition via Joint Learning

A highly efficient, fast, powerful and light-weight anime downloader and streamer for your favorite anime.

Adaptation through prediction: multisensory active inference torque control

Material for my PyConDE & PyData Berlin 2022 Talk "5 Steps to Speed Up Your Data-Analysis on a Single Core"

Pytorch-Swin-Unet-V2 - a modified version of Swin Unet based on Swin Transfomer V2

Pseudo-rng-app - whos needs science to make a random number when you have pseudoscience?

PyTorch Implementation for AAAI'21 "Do Response Selection Models Really Know What's Next? Utterance Manipulation Strategies for Multi-turn Response Selection"

Recurrent Neural Network Tutorial, Part 2 - Implementing a RNN in Python and Theano

Thermal Control of Laser Powder Bed Fusion using Deep Reinforcement Learning

hipCaffe: the HIP port of Caffe

Simple Tensorflow implementation of Toward Spatially Unbiased Generative Models (ICCV 2021)

Attention-driven Robot Manipulation (ARM) which includes Q-attention

Colossal-AI: A Unified Deep Learning System for Large-Scale Parallel Training

An original implementation of "MetaICL Learning to Learn In Context" by Sewon Min, Mike Lewis, Luke Zettlemoyer and Hannaneh Hajishirzi