[CVPR 2021] Counterfactual VQA: A Cause-Effect Look at Language Bias

Last update: Dec 03, 2022

Overview

Counterfactual VQA (CF-VQA)

This repository is the Pytorch implementation of our paper "Counterfactual VQA: A Cause-Effect Look at Language Bias" in CVPR 2021. This code is implemented as a fork of RUBi.

CF-VQA is proposed to capture and mitigate language bias in VQA from the view of causality. CF-VQA (1) captures the language bias as the direct causal effect of questions on answers, and (2) reduces the language bias by subtracting the direct language effect from the total causal effect.

If you find this paper helps your research, please kindly consider citing our paper in your publications.

@inproceedings{niu2020counterfactual,
  title={Counterfactual VQA: A Cause-Effect Look at Language Bias},
  author={Niu, Yulei and Tang, Kaihua and Zhang, Hanwang and Lu, Zhiwu and Hua, Xian-Sheng and Wen, Ji-Rong},
  booktitle={Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition},
  year={2021}
}

Summary

Installation
- Setup and dependencies
- Download datasets
Quick start
- Train a model
- Evaluate a model
Useful commands
Acknowledgment

Installation

1. Setup and dependencies

Install Anaconda or Miniconda distribution based on Python3+ from their downloads' site.

conda create --name cfvqa python=3.7
source activate cfvqa
pip install -r requirements.txt

2. Download datasets

Download annotations, images and features for VQA experiments:

bash cfvqa/datasets/scripts/download_vqa2.sh
bash cfvqa/datasets/scripts/download_vqacp2.sh

Quick start

Train a model

The boostrap/run.py file load the options contained in a yaml file, create the corresponding experiment directory and start the training procedure. For instance, you can train our best model on VQA-CP v2 (CFVQA+SUM+SMRL) by running:

python -m bootstrap.run -o cfvqa/options/vqacp2/smrl_cfvqa_sum.yaml

Then, several files are going to be created in logs/vqacp2/smrl_cfvqa_sum/:

[options.yaml] (copy of options)
[logs.txt] (history of print)
[logs.json] (batchs and epochs statistics)
[_vq_val_oe.json] (statistics for the language-prior based strategy, e.g., RUBi)
[_cfvqa_val_oe.json] (statistics for CF-VQA)
[_q_val_oe.json] (statistics for language-only branch)
[_v_val_oe.json] (statistics for vision-only branch)
[_all_val_oe.json] (statistics for the ensembled branch)
ckpt_last_engine.pth.tar (checkpoints of last epoch)
ckpt_last_model.pth.tar
ckpt_last_optimizer.pth.tar

Many options are available in the options directory. CFVQA represents the complete causal graph while cfvqas represents the simplified causal graph.

Evaluate a model

There is no test set on VQA-CP v2, our main dataset. The evaluation is done on the validation set. For a model trained on VQA v2, you can evaluate your model on the test set. In this example, boostrap/run.py load the options from your experiment directory, resume the best checkpoint on the validation set and start an evaluation on the testing set instead of the validation set while skipping the training set (train_split is empty). Thanks to --misc.logs_name, the logs will be written in the new logs_predicate.txt and logs_predicate.json files, instead of being appended to the logs.txt and logs.json files.

python -m bootstrap.run \
-o ./logs/vqacp2/smrl_cfvqa_sum/options.yaml \
--exp.resume last \
--dataset.train_split ''\
--dataset.eval_split val \
--misc.logs_name test

Useful commands

Use a specific GPU

For a specific experiment:

CUDA_VISIBLE_DEVICES=0 python -m boostrap.run -o cfvqa/options/vqacp2/smrl_cfvqa_sum.yaml

For the current terminal session:

export CUDA_VISIBLE_DEVICES=0

Overwrite an option

The boostrap.pytorch framework makes it easy to overwrite a hyperparameter. In this example, we run an experiment with a non-default learning rate. Thus, I also overwrite the experiment directory path:

python -m bootstrap.run -o cfvqa/options/vqacp2/smrl_cfvqa_sum.yaml \
--optimizer.lr 0.0003 \
--exp.dir logs/vqacp2/smrl_cfvqa_sum_lr,0.0003

Resume training

If a problem occurs, it is easy to resume the last epoch by specifying the options file from the experiment directory while overwritting the exp.resume option (default is None):

python -m bootstrap.run -o logs/vqacp2/smrl_cfvqa_sum/options.yaml \
--exp.resume last

Acknowledgment

Special thanks to the authors of RUBi, BLOCK, and bootstrap.pytorch, and the datasets used in this research project.

[CVPR 2021] Counterfactual VQA: A Cause-Effect Look at Language Bias

Related tags

Overview

Counterfactual VQA (CF-VQA)

Summary

Installation

1. Setup and dependencies

2. Download datasets

Quick start

Train a model

Evaluate a model

Useful commands

Use a specific GPU

Overwrite an option

Resume training

Acknowledgment

Owner

Yulei Niu

PyTorch version of the paper 'Enhanced Deep Residual Networks for Single Image Super-Resolution' (CVPRW 2017)

GNN-based Recommendation Benchmark

This code provides various models combining dilated convolutions with residual networks

Package for working with hypernetworks in PyTorch.

FCOS: Fully Convolutional One-Stage Object Detection (ICCV'19)

SigOpt wrappers for scikit-learn methods

Deep ViT Features as Dense Visual Descriptors

This repository is for EMNLP 2021 paper: It is Not as Good as You Think! Evaluating Simultaneous Machine Translation on Interpretation Data

The official implementation of NeurIPS 2021 paper: Finding Optimal Tangent Points for Reducing Distortions of Hard-label Attacks

WORD: Revisiting Organs Segmentation in the Whole Abdominal Region

System Combination for Grammatical Error Correction Based on Integer Programming

PyTorch implementation of Value Iteration Networks (VIN): Clean, Simple and Modular. Visualization in Visdom.

Everything you want about DP-Based Federated Learning, including Papers and Code. (Mechanism: Laplace or Gaussian, Dataset: femnist, shakespeare, mnist, cifar-10 and fashion-mnist. )

Predict multi paths to a moving person depending on his trajectory history.

Hardware accelerated, batchable and differentiable optimizers in JAX.

Scikit-learn compatible estimation of general graphical models

Multivariate Boosted TRee

Creative Applications of Deep Learning w/ Tensorflow

Microsoft Cognitive Toolkit (CNTK), an open source deep-learning toolkit

FinRL-Meta: A Universe for Data-Driven Financial Reinforcement Learning. 🔥

[CVPR 2021] Counterfactual VQA: A Cause-Effect Look at Language Bias

Related tags

Overview

Counterfactual VQA (CF-VQA)

Summary

Installation

1. Setup and dependencies

2. Download datasets

Quick start

Train a model

Evaluate a model

Useful commands

Use a specific GPU

Overwrite an option

Resume training

Acknowledgment

Owner

Yulei Niu

PyTorch version of the paper 'Enhanced Deep Residual Networks for Single Image Super-Resolution' (CVPRW 2017)

GNN-based Recommendation Benchmark

This code provides various models combining dilated convolutions with residual networks

Package for working with hypernetworks in PyTorch.

FCOS: Fully Convolutional One-Stage Object Detection (ICCV'19)

SigOpt wrappers for scikit-learn methods

Deep ViT Features as Dense Visual Descriptors

This repository is for EMNLP 2021 paper: It is Not as Good as You Think! Evaluating Simultaneous Machine Translation on Interpretation Data

The official implementation of NeurIPS 2021 paper: Finding Optimal Tangent Points for Reducing Distortions of Hard-label Attacks

WORD: Revisiting Organs Segmentation in the Whole Abdominal Region

System Combination for Grammatical Error Correction Based on Integer Programming

PyTorch implementation of Value Iteration Networks (VIN): Clean, Simple and Modular. Visualization in Visdom.

Everything you want about DP-Based Federated Learning, including Papers and Code. (Mechanism: Laplace or Gaussian, Dataset: femnist, shakespeare, mnist, cifar-10 and fashion-mnist. )

Predict multi paths to a moving person depending on his trajectory history.

Hardware accelerated, batchable and differentiable optimizers in JAX.

Scikit-learn compatible estimation of general graphical models

Multivariate Boosted TRee

Creative Applications of Deep Learning w/ Tensorflow

Microsoft Cognitive Toolkit (CNTK), an open source deep-learning toolkit

FinRL­-Meta: A Universe for Data­-Driven Financial Reinforcement Learning. 🔥

FinRL-Meta: A Universe for Data-Driven Financial Reinforcement Learning. 🔥