QR Kod

Semantic Alignment on Action for Image Captioning

Image captioning is a popular task in vision and language processing, which aims to generate textual descriptions for images. Previously, it simply used image and text as input with self-attention to capture global dependencies. Recent research further uses objects detected from the input image, so-...

Ful tanımlama

Kaydedildi:
Detaylı Bibliyografya
Asıl Yazarlar: Da Huo, Marc A. Kastner, Takatsugu Hirayama, Takahiro Komamizu, Yasutomo Kawanishi, Ichiro Ide
Materyal Türü: Artigo
Dil:Inglês
Baskı/Yayın Bilgisi: IEEE 2025-01-01
Seri Bilgileri:IEEE Access
Konular:
Online Erişim:https://ieeexplore.ieee.org/document/11237113/
Etiketler: Etiketle
Etiket eklenmemiş, İlk siz ekleyin!