QR kód

DDFAV: Remote Sensing Large Vision Language Models Dataset and Evaluation Benchmark

With the rapid development of large visual language models (LVLMs) and multimodal large language models (MLLMs), these models have demonstrated strong performance in various multimodal tasks. However, alleviating the generation of hallucinations remains a key challenge in LVLMs research. For remote...

Celý popis

Uloženo v:
Podrobná bibliografie
Hlavní autoři: Haodong Li, Xiaofeng Zhang, Haicheng Qu
Médium: Artigo
Jazyk:Inglês
Vydáno: MDPI AG 2025-02-01
Edice:Remote Sensing
Témata:
On-line přístup:https://www.mdpi.com/2072-4292/17/4/719
Tagy: Přidat tag
Žádné tagy, Buďte první, kdo vytvoří štítek k tomuto záznamu!