QR code

DeepSeek Under Attack: An Analysis of Jailbreak Attacks and Prompt-Level Defenses

Large Language Models (LLMs) with reasoning capabilities (e.g., DeepSeek-R1) gained substantial research and industry interest. However, their novel reasoning features may introduce vulnerabilities, especially to specific jailbreak attacks that exploit weaknesses in safety alignment. Despite growing...

Volledige beschrijving

Bewaard in:
Bibliografische gegevens
Hoofdauteurs: Victor Takashi Hayashi, Milton Pedro Pagliuso Neto, Charles Christian Miers, Fernando Frota Redigolo, Reginaldo Arakaki, Marcos Antonio Simplicio
Formaat: Artigo
Taal:Inglês
Gepubliceerd in: IEEE 2026-01-01
Reeks:IEEE Access
Onderwerpen:
Online toegang:https://ieeexplore.ieee.org/document/11595604/
Tags: Voeg label toe
Geen labels, Wees de eerste die dit record labelt!