Class-dependent and cross-modal memory network considering sentimental features for video-based captioning
The video-based commonsense captioning task aims to add multiple commonsense descriptions to video captions to understand video content better. This paper aims to consider the importance of cross-modal mapping. We propose a combined framework called Class-dependent and Cross-modal Memory Network con...
Guardado en:
| Autores principales: | , , , |
|---|---|
| Formato: | Artigo |
| Lenguaje: | Inglês |
| Publicado: |
Frontiers Media S.A.
2023-02-01
|
| Colección: | Frontiers in Psychology |
| Materias: | |
| Acceso en línea: | https://www.frontiersin.org/articles/10.3389/fpsyg.2023.1124369/full |
| Etiquetas: |
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
