On the Limitations of Visual-Semantic Embedding Networks for Image-to-Text Information Retrieval
Visual-semantic embedding (VSE) networks create joint image–text representations to map images and texts in a shared embedding space to enable various information retrieval-related tasks, such as image–text retrieval, image captioning, and visual question answering. The most recent state-of-the-art...
Kaydedildi:
| Asıl Yazarlar: | , , |
|---|---|
| Materyal Türü: | Artigo |
| Dil: | Inglês |
| Baskı/Yayın Bilgisi: |
MDPI AG
2021-07-01
|
| Seri Bilgileri: | Journal of Imaging |
| Konular: | |
| Online Erişim: | https://www.mdpi.com/2313-433X/7/8/125 |
| Etiketler: |
Etiket eklenmemiş, İlk siz ekleyin!
|
