Codice QR

VSRI:Visual Semantic Relational Interactor for Image Caption

Image captioning is one of the key objectives of multimodal image understanding.This paper aims to generate detail-rich and accurate image caption.Currently,mainstream image captioning methods focus on the interrelationships between regions,but ignore the visual semantic relationships between region...

Descrizione completa

Salvato in:
Dettagli Bibliografici
Autore principale: LIU Jian, YAO Renyuan, GAO Nan, LIANG Ronghua, CHEN Peng
Natura: Artigo
Lingua:Chinês
Pubblicazione: Editorial office of Computer Science 2025-08-01
Serie:Jisuanji kexue
Soggetti:
Accesso online:https://www.jsjkx.com/fileup/1002-137X/PDF/1002-137X-2025-52-8-222.pdf
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!