Dynamic Co-attention Network for Visual Question Answering

Mapping Intimacies ◽

10.1109/iscmi53840.2021.9654812 ◽

2021 ◽

Author(s):

Doaa B. Ebaid ◽

Magda M. Madbouly ◽

Adel A. El-Zoghabi

Keyword(s):

Question Answering ◽

Attention Network ◽

Visual Question Answering

Download Full-text

Multi-tier attention network using term-weighted question features for Visual Question Answering

Image and Vision Computing ◽

10.1016/j.imavis.2021.104291 ◽

2021 ◽

pp. 104291

Author(s):

Sruthy Manmadhan ◽

Binsu C. Kovoor

Keyword(s):

Question Answering ◽

Attention Network ◽

Visual Question Answering

Download Full-text

Multi-Modality Global Fusion Attention Network for Visual Question Answering

Electronics ◽

10.3390/electronics9111882 ◽

2020 ◽

Vol 9 (11) ◽

pp. 1882

Author(s):

Cheng Yang ◽

Weijia Wu ◽

Yuxing Wang ◽

Hong Zhou

Keyword(s):

Correct Answer ◽

Question Answering ◽

Human Beings ◽

Single Model ◽

Attention Network ◽

Attention Model ◽

Global Perspectives ◽

Visual Question Answering ◽

Previous State ◽

Visual question answering (VQA) requires a high-level understanding of both questions and images, along with visual reasoning to predict the correct answer. Therefore, it is important to design an effective attention model to associate key regions in an image with key words in a question. Up to now, most attention-based approaches only model the relationships between individual regions in an image and words in a question. It is not enough to predict the correct answer for VQA, as human beings always think in terms of global information, not only local information. In this paper, we propose a novel multi-modality global fusion attention network (MGFAN) consisting of stacked global fusion attention (GFA) blocks, which can capture information from global perspectives. Our proposed method computes co-attention and self-attention at the same time, rather than computing them individually. We validate our proposed method on the two most commonly used benchmarks, the VQA-v2 datasets. Experimental results show that the proposed method outperforms the previous state-of-the-art. Our best single model achieves 70.67% accuracy on the test-dev set of VQA-v2.

Download Full-text

Text-Guided Dual-Branch Attention Network for Visual Question Answering

Advances in Multimedia Information Processing – PCM 2018 - Lecture Notes in Computer Science ◽

10.1007/978-3-030-00764-5_69 ◽

2018 ◽

pp. 750-760 ◽

Author(s):

Mengfei Li ◽

Li Gu ◽

Yi Ji ◽

Chunping Liu

Keyword(s):

Question Answering ◽

Attention Network ◽

Visual Question Answering

Download Full-text

Research On Visual Question Answering Based On Deep Stacked Attention Network

Journal of Physics Conference Series ◽

10.1088/1742-6596/1873/1/012047 ◽

2021 ◽

Vol 1873 (1) ◽

pp. 012047

Author(s):

Zhu Xiaoqing ◽

Han Junjun

Keyword(s):

Question Answering ◽

Attention Network ◽

Visual Question Answering

Download Full-text

Co-Attention Network With Question Type for Visual Question Answering

IEEE Access ◽

10.1109/access.2019.2908035 ◽

2019 ◽

Vol 7 ◽

pp. 40771-40781 ◽

Author(s):

Chao Yang ◽

Mengqi Jiang ◽

Bin Jiang ◽

Weixin Zhou ◽

Keqin Li

Keyword(s):

Question Answering ◽

Question Type ◽

Attention Network ◽

Visual Question Answering

Download Full-text

Two-Step Joint Attention Network for Visual Question Answering

2017 3rd International Conference on Big Data Computing and Communications (BIGCOM) ◽

10.1109/bigcom.2017.17 ◽

2017 ◽

Author(s):

Weiming Zhang ◽

Chunhong Zhang ◽

Pei Liu ◽

Zhiqiang Zhan ◽

Xiaofeng Qiu

Keyword(s):

Joint Attention ◽

Question Answering ◽

Attention Network ◽

Visual Question Answering

Download Full-text

Word-to-region attention network for visual question answering

Multimedia Tools and Applications ◽

10.1007/s11042-018-6389-3 ◽

2018 ◽

Vol 78 (3) ◽

pp. 3843-3858 ◽

Author(s):

Liang Peng ◽

Yang Yang ◽

Yi Bin ◽

Ning Xie ◽

Fumin Shen ◽

...

Keyword(s):

Question Answering ◽

Attention Network ◽

Visual Question Answering

Download Full-text

Multi-Channel Co-Attention Network for Visual Question Answering

2020 International Joint Conference on Neural Networks (IJCNN) ◽

10.1109/ijcnn48605.2020.9207058 ◽

2020 ◽

Author(s):

Weidong Tian ◽

Bin He ◽

Nanxun Wang ◽

Zhongqiu Zhao

Keyword(s):

Question Answering ◽

Attention Network ◽

Visual Question Answering

Download Full-text

Relation-Aware Graph Attention Network for Visual Question Answering

2019 IEEE/CVF International Conference on Computer Vision (ICCV) ◽

10.1109/iccv.2019.01041 ◽

2019 ◽

Author(s):

Linjie Li ◽

Zhe Gan ◽

Yu Cheng ◽

Jingjing Liu

Keyword(s):

Question Answering ◽

Attention Network ◽

Visual Question Answering

Download Full-text

Triple attention network for sentimental visual question answering

Computer Vision and Image Understanding ◽

10.1016/j.cviu.2019.102829 ◽

2019 ◽

Vol 189 ◽

pp. 102829

Author(s):

Nelson Ruwa ◽

Qirong Mao ◽

Heping Song ◽

Hongjie Jia ◽

Ming Dong

Keyword(s):

Question Answering ◽

Attention Network ◽

Visual Question Answering

Download Full-text