Adversarial VQA: A New Benchmark for Evaluating the Robustness of VQA Models

Li, Linjie; Lei, Jie; Gan, Zhe; Liu, Jingjing

Full-text links:

Download:

Current browse context:

cs.CV

< prev | next >

new | recent | 2106

Computer Science > Computer Vision and Pattern Recognition

Title: Adversarial VQA: A New Benchmark for Evaluating the Robustness of VQA Models

Authors: Linjie Li, Jie Lei, Zhe Gan, Jingjing Liu

(Submitted on 1 Jun 2021 (v1), last revised 13 Aug 2021 (this version, v2))

Abstract: Benefiting from large-scale pre-training, we have witnessed significant performance boost on the popular Visual Question Answering (VQA) task. Despite rapid progress, it remains unclear whether these state-of-the-art (SOTA) models are robust when encountering examples in the wild. To study this, we introduce Adversarial VQA, a new large-scale VQA benchmark, collected iteratively via an adversarial human-and-model-in-the-loop procedure. Through this new benchmark, we discover several interesting findings. (i) Surprisingly, we find that during dataset collection, non-expert annotators can easily attack SOTA VQA models successfully. (ii) Both large-scale pre-trained models and adversarial training methods achieve far worse performance on the new benchmark than over standard VQA v2 dataset, revealing the fragility of these models while demonstrating the effectiveness of our adversarial dataset. (iii) When used for data augmentation, our dataset can effectively boost model performance on other robust VQA benchmarks. We hope our Adversarial VQA dataset can shed new light on robustness study in the community and serve as a valuable benchmark for future work.

Comments:	To appear in ICCV 2021; Website: this https URL
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Computation and Language (cs.CL)
Cite as:	arXiv:2106.00245 [cs.CV]
	(or arXiv:2106.00245v2 [cs.CV] for this version)

Submission history

From: Linjie Li [view email]
[v1] Tue, 1 Jun 2021 05:54:41 GMT (15704kb,D)
[v2] Fri, 13 Aug 2021 07:01:48 GMT (16801kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2106.00245v2

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computer Vision and Pattern Recognition

Title: Adversarial VQA: A New Benchmark for Evaluating the Robustness of VQA Models

Submission history