Código QR (código de barras bidimensional)

VSRI:Visual Semantic Relational Interactor for Image Caption

Image captioning is one of the key objectives of multimodal image understanding.This paper aims to generate detail-rich and accurate image caption.Currently,mainstream image captioning methods focus on the interrelationships between regions,but ignore the visual semantic relationships between region...

ver descrição completa

Na minha lista:
Detalhes bibliográficos
Autor principal: LIU Jian, YAO Renyuan, GAO Nan, LIANG Ronghua, CHEN Peng
Formato: Artigo
Idioma:Chinês
Publicado em: Editorial office of Computer Science 2025-08-01
coleção:Jisuanji kexue
Assuntos:
Acesso em linha:https://www.jsjkx.com/fileup/1002-137X/PDF/1002-137X-2025-52-8-222.pdf
Tags: Adicionar Tag
Sem tags, seja o primeiro a adicionar uma tag!