QR-koodi

PTCR: Knowledge-Based Visual Question Answering Framework Based on Large Language Model

Aiming at the problems of insufficient model input information and poor reasoning performance in knowledge-based visual question answering (VQA), this paper constructs a PTCR knowledge-based VQA framework based on large language model (LLM), which consists of four parts: answer candidate generation,...

Täydet tiedot

Tallennettuna:
Bibliografiset tiedot
Päätekijä: XUE Di, LI Xin, LIU Mingshuai
Aineistotyyppi: Artigo
Kieli:Chinês
Julkaistu: Journal of Computer Engineering and Applications Beijing Co., Ltd., Science Press 2024-11-01
Sarja:Jisuanji kexue yu tansuo
Aiheet:
Linkit:http://fcst.ceaj.org/fileup/1673-9418/PDF/2406028.pdf
Tagit: Lisää tagi
Ei tageja, Lisää ensimmäinen tagi!