Quantifying modality imbalance and visual jailbreak robustness in LLaVA via projected gradient descent
Abstract While Large Vision Language Models (LVLMs) exhibit remarkable capabilities, their visual modality introduces a critical attack surface that can bypass text only safety alignments. This paper evaluates the vulnerability of LLaVA-1.5 to targeted adversarial visual prompts designed to induce m...
محفوظ في:
| المؤلفون الرئيسيون: | , , |
|---|---|
| التنسيق: | Artigo |
| اللغة: | Inglês |
| منشور في: |
Springer
2026-05-01
|
| سلاسل: | Discover Applied Sciences |
| الموضوعات: | |
| الوصول للمادة أونلاين: | https://doi.org/10.1007/s42452-026-08793-w |
| الوسوم: |
لا توجد وسوم, كن أول من يضع وسما على هذه التسجيلة!
|
