QR code

Image-Caption Model Based on Fusion Feature

The encoder–decoder framework is the main frame of image captioning. The convolutional neural network (CNN) is usually used to extract grid-level features of the image, and the graph convolutional neural network (GCN) is used to extract the image’s region-level features. Grid-level features are poor...

Volledige beschrijving

Bewaard in:
Bibliografische gegevens
Hoofdauteurs: Yaogang Geng, Hongyan Mei, Xiaorong Xue, Xing Zhang
Formaat: Artigo
Taal:Inglês
Gepubliceerd in: MDPI AG 2022-09-01
Reeks:Applied Sciences
Onderwerpen:
Online toegang:https://www.mdpi.com/2076-3417/12/19/9861
Tags: Voeg label toe
Geen labels, Wees de eerste die dit record labelt!