Field notes.
Building Doberman in the open: AI agent security, incidents worth learning from, and the case for putting runtime authorization on the execution path.
-
8 min read
Pacing the Frontier Is Not Enough
Dario Amodei’s We Must Pace the Frontier makes a strong case that AI capability is beginning to move faster than our ability to understand and control it.
-
4 min read
When my guardrail crashes, the agent gets a no
The easiest way past most security checks is to make them fall over. I built Doberman so that its own failure is a denial, and that choice has a cost worth being honest about.
-
5 min read
What it cost me to cut the third leg of the lethal trifecta
Simon Willison’s rule says an agent can’t safely hold private data, read untrusted content, and have a way out. A coding agent needs all three, so I put a check on the third one. This is the bill.
-
6 min read
Approval Fatigue is a security issue, not a UX problem
I’ve spent the last two weeks making Doberman ask for approval less often. This increased safety and here’s why.
-
4 min read
An honest comparison: Cisco’s DefenseClaw and Doberman
Cisco's answer to runtime AI defense — what DefenseClaw gets right, and where a dedicated authorization layer makes a different bet.
-
21 min read
The AI Security Playbook
Takeaways from RAND's Practical Guide for Securing AI Models — and why AI security is shifting from model safety to execution security.
-
11 min read
When AI Output Stops Being Text
The Hugging Face incident shows why agent security cannot end with better prompts, safer models or stronger sandboxes.
Prefer email? These posts are mirrored on Substack ↗ — or subscribe to the RSS feed.