QR koda

Combining Region-Guided Attention and Attribute Prediction for Thangka Image Captioning Method

To enhance the understanding of the core regions in Thangka images and improve the richness of generated content during decoding, we propose a Thangka image captioning method based on Region-Guided Feature Enhancement and Attribute Prediction (RGFEAP). The image feature enhancement encoder, guided b...

Popoln opis

Shranjeno v:
Bibliografske podrobnosti
Principais autores: Fujun Zhang, Wendong Kang, Wenjin Hu
Format: Artigo
Jezik:Inglês
Izdano: IEEE 2025-01-01
Serija:IEEE Access
Teme:
Online dostop:https://ieeexplore.ieee.org/document/10833628/
Oznake: Označite
Brez oznak, prvi označite!