DeepSeek Under Attack: An Analysis of Jailbreak Attacks and Prompt-Level Defenses
Large Language Models (LLMs) with reasoning capabilities (e.g., DeepSeek-R1) gained substantial research and industry interest. However, their novel reasoning features may introduce vulnerabilities, especially to specific jailbreak attacks that exploit weaknesses in safety alignment. Despite growing...
Bewaard in:
| Hoofdauteurs: | , , , , , |
|---|---|
| Formaat: | Artigo |
| Taal: | Inglês |
| Gepubliceerd in: |
IEEE
2026-01-01
|
| Reeks: | IEEE Access |
| Onderwerpen: | |
| Online toegang: | https://ieeexplore.ieee.org/document/11595604/ |
| Tags: |
Geen labels, Wees de eerste die dit record labelt!
|
