The prompt injection series
Listed inPrompt Injection AttacksSafety & Ethicson
Years of worked examples showing why filtering does not solve injection, and why the dual-LLM pattern is the closest thing to a fix.
The failure modes that make headlines, and the controls that prevent them.
12 articles
Listed inPrompt Injection AttacksSafety & Ethicson
Years of worked examples showing why filtering does not solve injection, and why the dual-LLM pattern is the closest thing to a fix.
Listed inSecurity and Privacy ConcernsSafety & Ethicson
The industry checklist: injection, insecure output handling, supply chain, data leakage, and the rest — with mitigations.
Listed inConducting Adversarial TestingSafety & Ethicson
Red-teaming your own system before someone else does.
Listed inContent Moderation APIsSafety & Ethicson
Classifying input and output before either reaches a user.
Listed inBias and FairnessSafety & Ethicson
Where training data bias surfaces in product behaviour, and how to measure it.
Listed inConstraining Inputs and OutputsSafety & Ethicson
Allowlists, schemas, and length limits as a safety mechanism.
Listed inAI Safety and EthicsSafety & Ethicson
The obligations that come with shipping a system you cannot fully predict.
Listed inRobust Prompt EngineeringSafety & Ethicson
Instructions that survive hostile input and long conversations.
Listed inAdding End-User IDsSafety & Ethicson
Attributing requests so providers can flag abuse and you can trace it.
Listed inSecurity and Privacy ConcernsSafety & Ethicson
PII in prompts, data retention, and what leaves your perimeter.
Listed inPrompt Injection AttacksSafety & Ethicson
Untrusted text that rewrites your instructions, and why filtering isn't enough.
Listed inKnow Your Customers & Use CasesSafety & Ethicson
Scoping who the system serves so you can bound what it must handle.