Conferences >2017 IEEE International Confe...

Structured Attentions for Visual Question Answering

Download PDF
Download References
Request Permissions
Save to
Alerts

Abstract:

Visual attention, which assigns weights to image regions according to their relevance to a question, is considered as an indispensable part by most Visual Question Answer...Show More

Metadata

Abstract:

Visual attention, which assigns weights to image regions according to their relevance to a question, is considered as an indispensable part by most Visual Question Answering models. Although the questions may involve complex rela- tions among multiple regions, few attention models can ef- fectively encode such cross-region relations. In this paper, we demonstrate the importance of encoding such relations by showing the limited effective receptive field of ResNet on two datasets, and propose to model the visual attention as a multivariate distribution over a grid-structured Con- ditional Random Field on image regions. We demonstrate how to convert the iterative inference algorithms, Mean Field and Loopy Belief Propagation, as recurrent layers of an end-to-end neural network. We empirically evalu- ated our model on 3 datasets, in which it surpasses the best baseline model of the newly released CLEVR dataset [13] by 9.5%, and the best published model on the VQA dataset [3] by 1.25%. Source code is available at https://github.com/zhuchen03/vqa-sva.

Published in: 2017 IEEE International Conference on Computer Vision (ICCV)

Date of Conference: 22-29 October 2017

Date Added to IEEE Xplore: 25 December 2017

ISBN Information:

Electronic ISSN: 2380-7504

DOI: 10.1109/ICCV.2017.145

Conference Location: Venice, Italy

Contents

References is not available for this document.

Structured Attentions for Visual Question Answering

Abstract:

Metadata

Abstract:

References

IEEE Account

Purchase Details

Profile Information

Need Help?

Structured Attentions for Visual Question Answering

Alerts

Abstract:

Metadata

Abstract:

Authors

Figures

References

Citations

Keywords

Metrics

Footnotes

References

IEEE Account

Purchase Details

Profile Information

Need Help?