Skip to main content
Single turn attacks are prompt injection techniques that attempt to manipulate the model in a single message exchange. These are the most common type of prompt injection attacks.

Attack Categories

Jailbreaking

Bypass safety measures through persona adoption and role manipulation

Encoding & Obfuscation

Hide malicious content using encoding and obfuscation techniques

Structural

Exploit input structure and format to bypass filters

Language-Based

Use language variations to evade detection

Jailbreaking Techniques

Direct attempts to bypass model safety measures through persona adoption and instruction manipulation.

Encoding & Obfuscation

Attacks that hide malicious content using various encoding and obfuscation techniques.

Structural Attacks

Attacks that exploit input structure or format to bypass content filters.

Language-Based Attacks

Attacks that use language variations to evade detection.

Multimodal Attacks


Quick Start Example