Analysis and field notes on AI security: prompt injection, agentic AI risk, interpretability and the frameworks shaping how organisations protect AI.
The old question was whether our AI is secure. The one that matters now is whether we can trust it to act.
Read the article
1YouSet the goal, not every step.2The agentPlans and acts across your tools.3A checkpoint on every actionEach action is checked against the goal.4Held for approvalAn unusual payment waits for a person.