V-PROM: A Benchmark for Visual Reasoning Using Visual Progressive Matrices

Teney, Damien; Wang, Peng; Cao, Jiewei; Liu, Lingqiao; Shen, Chunhua; Hengel, Anton van den

Computer Science > Computer Vision and Pattern Recognition

arXiv:1907.12271 (cs)

[Submitted on 29 Jul 2019]

Title:V-PROM: A Benchmark for Visual Reasoning Using Visual Progressive Matrices

Authors:Damien Teney, Peng Wang, Jiewei Cao, Lingqiao Liu, Chunhua Shen, Anton van den Hengel

View PDF

Abstract:One of the primary challenges faced by deep learning is the degree to which current methods exploit superficial statistics and dataset bias, rather than learning to generalise over the specific representations they have experienced. This is a critical concern because generalisation enables robust reasoning over unseen data, whereas leveraging superficial statistics is fragile to even small changes in data distribution. To illuminate the issue and drive progress towards a solution, we propose a test that explicitly evaluates abstract reasoning over visual data. We introduce a large-scale benchmark of visual questions that involve operations fundamental to many high-level vision tasks, such as comparisons of counts and logical operations on complex visual properties. The benchmark directly measures a method's ability to infer high-level relationships and to generalise them over image-based concepts. It includes multiple training/test splits that require controlled levels of generalization. We evaluate a range of deep learning architectures, and find that existing models, including those popular for vision-and-language tasks, are unable to solve seemingly-simple instances. Models using relational networks fare better but leave substantial room for improvement.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1907.12271 [cs.CV]
	(or arXiv:1907.12271v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1907.12271

Submission history

From: Damien Teney [view email]
[v1] Mon, 29 Jul 2019 08:28:33 UTC (3,474 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:V-PROM: A Benchmark for Visual Reasoning Using Visual Progressive Matrices

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:V-PROM: A Benchmark for Visual Reasoning Using Visual Progressive Matrices

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators