Adversarial Regularization for Visual Question Answering: Strengths, Shortcomings, and Side Effects

Grand, Gabriel; Belinkov, Yonatan

Computer Science > Machine Learning

arXiv:1906.08430 (cs)

[Submitted on 20 Jun 2019]

Title:Adversarial Regularization for Visual Question Answering: Strengths, Shortcomings, and Side Effects

Authors:Gabriel Grand, Yonatan Belinkov

View PDF

Abstract:Visual question answering (VQA) models have been shown to over-rely on linguistic biases in VQA datasets, answering questions "blindly" without considering visual context. Adversarial regularization (AdvReg) aims to address this issue via an adversary sub-network that encourages the main model to learn a bias-free representation of the question. In this work, we investigate the strengths and shortcomings of AdvReg with the goal of better understanding how it affects inference in VQA models. Despite achieving a new state-of-the-art on VQA-CP, we find that AdvReg yields several undesirable side-effects, including unstable gradients and sharply reduced performance on in-domain examples. We demonstrate that gradual introduction of regularization during training helps to alleviate, but not completely solve, these issues. Through error analyses, we observe that AdvReg improves generalization to binary questions, but impairs performance on questions with heterogeneous answer distributions. Qualitatively, we also find that regularized models tend to over-rely on visual features, while ignoring important linguistic cues in the question. Our results suggest that AdvReg requires further refinement before it can be considered a viable bias mitigation technique for VQA.

Comments:	In Proceedings of the 2nd Workshop on Shortcomings in Vision and Language (SiVL) at NAACL-HLT 2019
Subjects:	Machine Learning (cs.LG); Computation and Language (cs.CL); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (stat.ML)
Cite as:	arXiv:1906.08430 [cs.LG]
	(or arXiv:1906.08430v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1906.08430

Submission history

From: Gabriel Grand [view email]
[v1] Thu, 20 Jun 2019 03:28:09 UTC (7,997 KB)

Computer Science > Machine Learning

Title:Adversarial Regularization for Visual Question Answering: Strengths, Shortcomings, and Side Effects

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Adversarial Regularization for Visual Question Answering: Strengths, Shortcomings, and Side Effects

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators