Código QR

Semantic Alignment on Action for Image Captioning

Image captioning is a popular task in vision and language processing, which aims to generate textual descriptions for images. Previously, it simply used image and text as input with self-attention to capture global dependencies. Recent research further uses objects detected from the input image, so-...

ver descrição completa

Na minha lista:
Detalhes bibliográficos
Principais autores: Da Huo, Marc A. Kastner, Takatsugu Hirayama, Takahiro Komamizu, Yasutomo Kawanishi, Ichiro Ide
Formato: Artigo
Idioma:Inglês
Publicado em: IEEE 2025-01-01
Colecção:IEEE Access
Assuntos:
Acesso em linha:https://ieeexplore.ieee.org/document/11237113/
Tags: Adicionar Tag
Sem tags, seja o primeiro a adicionar uma tag!