Skip to content

Opening book details…

Can I read Image-Question Co-Attention for VQA on EtoBox?

Image-Question Co-Attention for VQA by deelipvenkat is a document available to read on EtoBox.

What is Image-Question Co-Attention for VQA about?

This document presents a project focused on Visual Question Answering (VQA) using an Image-Question-Linguistic Co-Attention model to improve accuracy in answering multiple choice questions about images. The authors developed an improved Multi-Layer Perceptron (MLP) model that incorporates GloVe embeddings, ResNet image features, and co-attention mechanisms, achieving significant performance enhancements on the Visual7W dataset. The study also explores the integration of Part-of-Speech (POS) tagging to furth

Author
deelipvenkat
Language
EN