QR-Code

Lightweight visual accessibility LLaVA architecture

Abstract At present, many models focus on the field of blind barrier-free recognition, but most of them are simple visual recognition. Multimodal visual language models face the challenges of high computational cost and poor real-time performance in blind barrier-free recognition tasks, which limits...

Ausführliche Beschreibung

Gespeichert in:
Bibliografische Detailangaben
Hauptverfasser: Zhiyin Han, Xiaoqun Liu, Juan Hao
Format: Artigo
Sprache:Inglês
Veröffentlicht: Nature Portfolio 2025-11-01
Schriftenreihe:Scientific Reports
Schlagworte:
Online-Zugang:https://doi.org/10.1038/s41598-025-23023-w
Tags: Tag hinzufügen
Keine Tags, Fügen Sie das erste Tag hinzu!