VSRI:Visual Semantic Relational Interactor for Image Caption
Image captioning is one of the key objectives of multimodal image understanding.This paper aims to generate detail-rich and accurate image caption.Currently,mainstream image captioning methods focus on the interrelationships between regions,but ignore the visual semantic relationships between region...
Salvato in:
| Autore principale: | |
|---|---|
| Natura: | Artigo |
| Lingua: | Chinês |
| Pubblicazione: |
Editorial office of Computer Science
2025-08-01
|
| Serie: | Jisuanji kexue |
| Soggetti: | |
| Accesso online: | https://www.jsjkx.com/fileup/1002-137X/PDF/1002-137X-2025-52-8-222.pdf |
| Tags: |
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
