QR رمز

Quantifying modality imbalance and visual jailbreak robustness in LLaVA via projected gradient descent

Abstract While Large Vision Language Models (LVLMs) exhibit remarkable capabilities, their visual modality introduces a critical attack surface that can bypass text only safety alignments. This paper evaluates the vulnerability of LLaVA-1.5 to targeted adversarial visual prompts designed to induce m...

وصف كامل

محفوظ في:
التفاصيل البيبلوغرافية
المؤلفون الرئيسيون: Saklain Abdullah, Riad Hossain, Mahfuzulhoq Chowdhury
التنسيق: Artigo
اللغة:Inglês
منشور في: Springer 2026-05-01
سلاسل:Discover Applied Sciences
الموضوعات:
الوصول للمادة أونلاين:https://doi.org/10.1007/s42452-026-08793-w
الوسوم: إضافة وسم
لا توجد وسوم, كن أول من يضع وسما على هذه التسجيلة!