Owen Castellanos

Covers agent security and incidents: what an attacker can reach once an AI system holds credentials.

Owen Castellanos writes The Guardrail coverage of agent security and incidents. He is interested in one question above the others: when a model is persuaded to do the wrong thing, what is it authorised to do next?

He writes post-incident analysis the way an engineer reads one, looking for the telemetry gap rather than the villain. Where an incident is described publicly, he says who reported it first and what remains unconfirmed.

He is sceptical of controls that depend on detecting an attack, and interested in controls that make a successful attack boring.

How to reach Owen

Tips, corrections and documents go to owen@theguardrailreport.com. If you are reporting an error in a published briefing, quote the sentence and, where it concerns a legal obligation, the provision you believe we misread.

Briefings by Owen Castellanos

OC

Owen Castellanos, Security Editor

Covers agent security and incidents: what an attacker can reach once an AI system holds credentials.