Co-Attention Network With Question Type for Visual Question Answering
Visual Question Answering (VQA) is a challenging multi-modal learning task since it requires an understanding of both visual and textual modalities simultaneously. Therefore, the approaches used to represent the images and questions in a fine-grained manner play key roles in the performance. In orde...
Збережено в:
| Автори: | , , , , |
|---|---|
| Формат: | Artigo |
| Мова: | Inglês |
| Опубліковано: |
IEEE
2019-01-01
|
| Серія: | IEEE Access |
| Предмети: | |
| Онлайн доступ: | https://ieeexplore.ieee.org/document/8676009/ |
| Теги: |
Немає тегів, Будьте першим, хто поставить тег для цього запису!
|
