QR կոդ

Improving visual question answering for remote sensing via alternate-guided attention and combined loss

Visual question answering (VQA) for remote sensing (RS) images offers a typical multi-modal task advanced by natural language processing and computer vision technologies. Still, it remains a problem subjected to two aspects of factors. First, the RS image contains a wealth of visual elements but is...

Ամբողջական նկարագրություն

Պահպանված է:
Մատենագիտական մանրամասներ
Հիմնական հեղինակներ: Jiangfan Feng, Etao Tang, Maimai Zeng, Zhujun Gu, Pinglang Kou, Wei Zheng
Ձևաչափ: Artigo
Լեզու:Inglês
Հրապարակվել է: Elsevier 2023-08-01
Շարք:International Journal of Applied Earth Observations and Geoinformation
Խորագրեր:
Առցանց հասանելիություն:http://www.sciencedirect.com/science/article/pii/S1569843223002510
Ցուցիչներ: Ավելացրեք ցուցիչ
Չկան պիտակներ, Եղեք առաջինը, ով նշում է այս գրառումը!