DSGEM: Dual scene graph enhancement module‐based visual question answering
Abstract Visual Question Answering (VQA) aims to appropriately answer a text question by understanding the image content. Attention‐based VQA models mine the implicit relationships between objects according to the feature similarity, which neglects the explicit relationships between objects, for exa...
שמור ב:
| Principais autores: | , , , , , |
|---|---|
| פורמט: | Artigo |
| שפה: | Inglês |
| יצא לאור: |
Wiley
2023-09-01
|
| סדרה: | IET Computer Vision |
| נושאים: | |
| גישה מקוונת: | https://doi.org/10.1049/cvi2.12186 |
| תגים: |
אין תגיות, היה/י הראשונ/ה לתייג את הרשומה!
|
