Lightweight visual accessibility LLaVA architecture
Abstract At present, many models focus on the field of blind barrier-free recognition, but most of them are simple visual recognition. Multimodal visual language models face the challenges of high computational cost and poor real-time performance in blind barrier-free recognition tasks, which limits...
Bewaard in:
| Hoofdauteurs: | , , |
|---|---|
| Formaat: | Artigo |
| Taal: | Inglês |
| Gepubliceerd in: |
Nature Portfolio
2025-11-01
|
| Reeks: | Scientific Reports |
| Onderwerpen: | |
| Online toegang: | https://doi.org/10.1038/s41598-025-23023-w |
| Tags: |
Geen labels, Wees de eerste die dit record labelt!
|
