All drills

Prompt Injection & Data Leakage

What you'll be able to do

Catch a RAG or agent architecture that treats retrieved content as safe context before a hostile document proves it isn't.

The security or trust-and-safety review that happens right before a RAG/agent system connects to any data source the team doesn't fully control.
Start this internship
Create an account to unlock the 5 sections, the workbench, and AskThili.
Begin

Sections

1. Prompt Injection & Data Leakage
🔒 locked
2. The Document Is Attacker-Controlled
🔒 locked
3. Where Leakage Actually Happens
🔒 locked
4. Work the Case
🔒 locked
5. The Memo
🔒 locked

Dig deeper

🔗The security risk memo template (template pack)
code

Part of these learning paths

I'm a leader and I want to evaluate and govern AI investments in my organization
View path →