Lightweight visual accessibility LLaVA architecture
Abstract At present, many models focus on the field of blind barrier-free recognition, but most of them are simple visual recognition. Multimodal visual language models face the challenges of high computational cost and poor real-time performance in blind barrier-free recognition tasks, which limits...
Gespeichert in:
| Hauptverfasser: | , , |
|---|---|
| Format: | Artigo |
| Sprache: | Inglês |
| Veröffentlicht: |
Nature Portfolio
2025-11-01
|
| Schriftenreihe: | Scientific Reports |
| Schlagworte: | |
| Online-Zugang: | https://doi.org/10.1038/s41598-025-23023-w |
| Tags: |
Keine Tags, Fügen Sie das erste Tag hinzu!
|
