Challenges and Prospects in Vision and Language Research

Kushal Kafle; Robik Shrestha; Christopher Kanan; Christopher Kanan; Christopher Kanan

doi:10.3389/frai.2019.00028

Frontiers in Artificial Intelligence (Dec 2019)

Challenges and Prospects in Vision and Language Research

Kushal Kafle,
Robik Shrestha,
Christopher Kanan,
Christopher Kanan,
Christopher Kanan

Affiliations

Kushal Kafle: Center for Imaging Science, Rochester Institute of Technology, Rochester, NY, United States
Robik Shrestha: Center for Imaging Science, Rochester Institute of Technology, Rochester, NY, United States
Christopher Kanan: Center for Imaging Science, Rochester Institute of Technology, Rochester, NY, United States
Christopher Kanan: Paige, New York, NY, United States
Christopher Kanan: Cornell Tech, New York, NY, United States

DOI: https://doi.org/10.3389/frai.2019.00028
Journal volume & issue: Vol. 2

Abstract

Read online

Language grounded image understanding tasks have often been proposed as a method for evaluating progress in artificial intelligence. Ideally, these tasks should test a plethora of capabilities that integrate computer vision, reasoning, and natural language understanding. However, the datasets and evaluation procedures used in these tasks are replete with flaws which allows the vision and language (V&L) algorithms to achieve a good performance without a robust understanding of vision and language. We argue for this position based on several recent studies in V&L literature and our own observations of dataset bias, robustness, and spurious correlations. Finally, we propose that several of these challenges can be mitigated by creation of carefully designed benchmarks.

Published in Frontiers in Artificial Intelligence

ISSN: 2624-8212 (Online)
Publisher: Frontiers Media S.A.
Country of publisher: Switzerland
LCC subjects: Science: Mathematics: Instruments and machines: Electronic computers. Computer science
Website: https://www.frontiersin.org/journals/artificial-intelligence#

About the journal

Abstract

Keywords