Código QR

Deconfounded fashion image captioning with transformer and multimodal retrieval

Background: The annotation of fashion images is a significantly important task in the fashion industry as well as social media and e-commerce. However, owing to the complexity and diversity of fashion images, this task entails multiple challenges, including the lack of fine-grained captions and conf...

Descrición completa

Gardado en:
Detalles Bibliográficos
Principais autores: Tao Peng, Weiqiao Yin, Junping Liu, Li Li, Xinrong Hu
Formato: Artigo
Idioma:Inglês
Publicado: KeAi Communications Co., Ltd. 2025-04-01
Series:Virtual Reality & Intelligent Hardware
Assuntos:
Acceso en liña:http://www.sciencedirect.com/science/article/pii/S2096579624000494
Tags: Engadir etiqueta
Sen Etiquetas, Sexa o primeiro en etiquetar este rexistro!