Towards Debiasing NLU Models from Unknown Biases

Last update: Jun 14, 2022

Overview

Towards Debiasing NLU Models from Unknown Biases

Abstract: NLU models often exploit biased features to achieve high dataset-specific performance without properly learning the intended task. Recently proposed debiasing methods are shown to be effective in mitigating this tendency. However, these methods rely on a major assumption that the type of biased features is known a-priori, which limits their application to many NLU tasks and datasets. In this work, we present the first step to bridge this gap by introducing a self-debiasing framework that prevents models from mainly utilizing biases without knowing them in advance. The proposed framework is general and complementary to the existing debiasing methods. We show that the proposed framework allows these existing methods to retain the improvement on the challenge datasets (i.e., sets of examples designed to expose models’ reliance to biases) without specifically targeting certain biases. Furthermore, the evaluation suggests that applying the framework results in improved overall robustness.

The repository contains the code to reproduce our work in debiasing NLU models without prior information on biases. We provide 3 runs of experiment that are shown in our paper:

Debias MNLI model from syntactic bias and evaluate on HANS as the out-of-distribution data using example reweighting.
Debias MNLI model from syntactic bias and evaluate on HANS as the out-of-distribution data using product of expert.
Debias MNLI model from syntactic bias and evaluate on HANS as the out-of-distribution data using confidence regularization.

Requirements

The code requires python >= 3.6 and pytorch >= 1.1.0.

Additional required dependencies can be found in requirements.txt. Install all requirements by running:

pip install -r requirements.txt

Data

Our experiments use MNLI dataset version provided by GLUE benchmark. Download the file from here, and unzip under the directory ./dataset The dataset directory should be structured as the following:

└── dataset 
    └── MNLI
        ├── train.tsv
        ├── dev_matched.tsv
        ├── dev_mismatched.tsv
        ├── dev_mismatched.tsv

Running the experiments

For each evaluation setting, use the --mode arguments to set the appropriate loss function. Choose the annealed version of the loss function for reproducing the annealed results.

To reproduce our result on MNLI ⮕ HANS, run the following:

cd src/
CUDA_VISIBLE_DEVICES=9 python train_distill_bert.py \
  --output_dir ../experiments_self_debias_mnli_seed111/bert_reweighted_sampled2K_teacher_seed111_annealed_1to08 \
  --do_train --do_eval --mode reweight_by_teacher_annealed \
  --custom_teacher ../teacher_preds/mnli_trained_on_sample2K_seed111.json --seed 111 --which_bias hans

Biased examples identification

To obtain predictions of the shallow models, we train the same model architecture on the fraction of the dataset. For MNLI we subsample 2000 examples and train the model for 5 epochs. For obtaining shallow models of other datasets please see the appendix of our paper. The shallow model can be obtained with the command below:

cd src/
CUDA_VISIBLE_DEVICES=9 python train_distill_bert.py \
 --output_dir ../experiments_shallow_mnli/bert_base_sampled2K_seed111 \
 --do_train --do_eval --do_eval_on_train --mode none\
 --seed 111 --which_bias hans --debug --num_train_epochs 5 --debug_num 2000

Once the training and the evaluation on train set is done, copy the probability json files in the output directory to ../teacher_preds/mnli_trained_on_sample2K_seed111.json.

Expected results

Results on the MNLI ⮕ HANS setting without annealing:

Mode	Seed	MNLI-m	MNLI-mm	HANS avg.
None	111	84.57	84.72	62.04
reweighting	111	81.8	82.3	72.1
PoE	111	81.5	81.1	70.3
conf-reg	222	83.7	84.1	68.7

Towards Debiasing NLU Models from Unknown Biases

Related tags

Overview

Towards Debiasing NLU Models from Unknown Biases

Requirements

Data

Running the experiments

Biased examples identification

Expected results

Owner

Ubiquitous Knowledge Processing Lab

unofficial pytorch implement of "Squareplus: A Softplus-Like Algebraic Rectifier"

Social Distancing Detector

Rule based classification A hotel s customers dataset

Garbage Detection system which will detect objects based on whether it is plastic waste or plastics or just garbage.

Use graph-based analysis to re-classify stocks and to improve Markowitz portfolio optimization

Official PyTorch implementation for FastDPM, a fast sampling algorithm for diffusion probabilistic models

A data-driven maritime port simulator

Federated Deep Reinforcement Learning for the Distributed Control of NextG Wireless Networks.

An evaluation toolkit for voice conversion models.

Face recognize and crop them

Individual Treatment Effect Estimation

MISSFormer: An Effective Medical Image Segmentation Transformer

Adjusting for Autocorrelated Errors in Neural Networks for Time Series

Towards the D-Optimal Online Experiment Design for Recommender Selection (KDD 2021)

Implementation of a protein autoregressive language model, but with autoregressive infilling objective (editing subsequences capability)

pytorch implementation of "Contrastive Multiview Coding", "Momentum Contrast for Unsupervised Visual Representation Learning", and "Unsupervised Feature Learning via Non-Parametric Instance-level Discrimination"

Progressive Growing of GANs for Improved Quality, Stability, and Variation

Learning Tracking Representations via Dual-Branch Fully Transformer Networks

Weighted QMIX: Expanding Monotonic Value Function Factorisation

Gans-in-action - Companion repository to GANs in Action: Deep learning with Generative Adversarial Networks