DeepSeek Under Attack: An Analysis of Jailbreak Attacks and Prompt-Level Defenses
Large Language Models (LLMs) with reasoning capabilities (e.g., DeepSeek-R1) gained substantial research and industry interest. However, their novel reasoning features may introduce vulnerabilities, especially to specific jailbreak attacks that exploit weaknesses in safety alignment. Despite growing...
Salvato in:
| Autori principali: | , , , , , |
|---|---|
| Natura: | Artigo |
| Lingua: | Inglês |
| Pubblicazione: |
IEEE
2026-01-01
|
| Serie: | IEEE Access |
| Soggetti: | |
| Accesso online: | https://ieeexplore.ieee.org/document/11595604/ |
| Tags: |
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
