VSRI:Visual Semantic Relational Interactor for Image Caption
Image captioning is one of the key objectives of multimodal image understanding.This paper aims to generate detail-rich and accurate image caption.Currently,mainstream image captioning methods focus on the interrelationships between regions,but ignore the visual semantic relationships between region...
Na minha lista:
| Autor principal: | |
|---|---|
| Formato: | Artigo |
| Idioma: | Chinês |
| Publicado em: |
Editorial office of Computer Science
2025-08-01
|
| coleção: | Jisuanji kexue |
| Assuntos: | |
| Acesso em linha: | https://www.jsjkx.com/fileup/1002-137X/PDF/1002-137X-2025-52-8-222.pdf |
| Tags: |
Sem tags, seja o primeiro a adicionar uma tag!
|
