কিউআর কোড

Deconfounded fashion image captioning with transformer and multimodal retrieval

Background: The annotation of fashion images is a significantly important task in the fashion industry as well as social media and e-commerce. However, owing to the complexity and diversity of fashion images, this task entails multiple challenges, including the lack of fine-grained captions and conf...

সম্পূর্ণ বিবরণ

সংরক্ষণ করুন:
গ্রন্থ-পঞ্জীর বিবরন
প্রধান লেখক: Tao Peng, Weiqiao Yin, Junping Liu, Li Li, Xinrong Hu
বিন্যাস: Artigo
ভাষা:Inglês
প্রকাশিত: KeAi Communications Co., Ltd. 2025-04-01
মালা:Virtual Reality & Intelligent Hardware
বিষয়গুলি:
অনলাইন ব্যবহার করুন:http://www.sciencedirect.com/science/article/pii/S2096579624000494
ট্যাগগুলো: ট্যাগ যুক্ত করুন
কোনো ট্যাগ নেই, প্রথমজন হিসাবে ট্যাগ করুন!