QR code

Lightweight visual accessibility LLaVA architecture

Abstract At present, many models focus on the field of blind barrier-free recognition, but most of them are simple visual recognition. Multimodal visual language models face the challenges of high computational cost and poor real-time performance in blind barrier-free recognition tasks, which limits...

Volledige beschrijving

Bewaard in:
Bibliografische gegevens
Hoofdauteurs: Zhiyin Han, Xiaoqun Liu, Juan Hao
Formaat: Artigo
Taal:Inglês
Gepubliceerd in: Nature Portfolio 2025-11-01
Reeks:Scientific Reports
Onderwerpen:
Online toegang:https://doi.org/10.1038/s41598-025-23023-w
Tags: Voeg label toe
Geen labels, Wees de eerste die dit record labelt!