Skip to main content

Learn how attackers manipulate AI models through malicious inputs, and how to stop them.

This module covers one of the most prevalent AI attack classes: prompt injection. Learners begin with direct injection techniques before moving on to jailbreaking and instruction smuggling via external content. The module then shifts to the defensive side, covering hardening techniques including filters, guards, and template isolation. Two challenge rooms provide hands-on red experience, reinforcing both attack recognition and mitigation strategies.

What are modules?

A learning pathway is made up of modules, and a module is made of bite-sized rooms (think of a room like a mini security lab).

Hierarchical diagram showing how learning pathways contain modules, which contain individual rooms.