QR Kod

On the Limitations of Visual-Semantic Embedding Networks for Image-to-Text Information Retrieval

Visual-semantic embedding (VSE) networks create joint image–text representations to map images and texts in a shared embedding space to enable various information retrieval-related tasks, such as image–text retrieval, image captioning, and visual question answering. The most recent state-of-the-art...

Ful tanımlama

Kaydedildi:
Detaylı Bibliyografya
Asıl Yazarlar: Yan Gong, Georgina Cosma, Hui Fang
Materyal Türü: Artigo
Dil:Inglês
Baskı/Yayın Bilgisi: MDPI AG 2021-07-01
Seri Bilgileri:Journal of Imaging
Konular:
Online Erişim:https://www.mdpi.com/2313-433X/7/8/125
Etiketler: Etiketle
Etiket eklenmemiş, İlk siz ekleyin!