QR код

Co-Attention Network With Question Type for Visual Question Answering

Visual Question Answering (VQA) is a challenging multi-modal learning task since it requires an understanding of both visual and textual modalities simultaneously. Therefore, the approaches used to represent the images and questions in a fine-grained manner play key roles in the performance. In orde...

Повний опис

Збережено в:
Бібліографічні деталі
Автори: Chao Yang, Mengqi Jiang, Bin Jiang, Weixin Zhou, Keqin Li
Формат: Artigo
Мова:Inglês
Опубліковано: IEEE 2019-01-01
Серія:IEEE Access
Предмети:
Онлайн доступ:https://ieeexplore.ieee.org/document/8676009/
Теги: Додати тег
Немає тегів, Будьте першим, хто поставить тег для цього запису!