Dark Promptery
A field reference · offensive-security research

Dark
Promptery

A catalogue of the techniques that get text past a language model's safety training, documented the way you would actually chain them. Every entry states the mechanism (what it is), why it slips the filter, a canonical example, and the detection or defense that neutralises it. Organised by attack layer, from single codepoints to agent-level injection.

v1.2.0 updated 2026-08-17 369 techniques CC BY 4.0 by SamsonCyber

The lay of the landHundreds of techniques, twenty categories, and they do not spread evenly. Where they pile up is where the model is softest, and where the cheapest transform buys the most coverage.

Fig 1 Distribution of documented techniques across the twenty categories, grouped by attack stage. The obfuscation and automated-attack families dominate the corpus, the surface where a small, mechanical transform still buys a large coverage gap. Click any bar to jump to its section.
// no techniques match that filter