Semantic Alignment on Action for Image Captioning
Image captioning is a popular task in vision and language processing, which aims to generate textual descriptions for images. Previously, it simply used image and text as input with self-attention to capture global dependencies. Recent research further uses objects detected from the input image, so-...
Kaydedildi:
| Asıl Yazarlar: | , , , , , |
|---|---|
| Materyal Türü: | Artigo |
| Dil: | Inglês |
| Baskı/Yayın Bilgisi: |
IEEE
2025-01-01
|
| Seri Bilgileri: | IEEE Access |
| Konular: | |
| Online Erişim: | https://ieeexplore.ieee.org/document/11237113/ |
| Etiketler: |
Etiket eklenmemiş, İlk siz ekleyin!
|
