Can I read Image-Question Co-Attention for VQA on EtoBox?
Image-Question Co-Attention for VQA by deelipvenkat is a document available to read on EtoBox.
What is Image-Question Co-Attention for VQA about?
This document presents a project focused on Visual Question Answering (VQA) using an Image-Question-Linguistic Co-Attention model to improve accuracy in answering multiple choice questions about images. The authors developed an improved Multi-Layer Perceptron (MLP) model that incorporates GloVe embeddings, ResNet image features, and co-attention mechanisms, achieving significant performance enhancements on the Visual7W dataset. The study also explores the integration of Part-of-Speech (POS) tagging to furth
- Author
- deelipvenkat
- Language
- EN