Run the agent in shadow mode before it changes anything
An agent’s preview is useful only if the boundary prevents writes. Review the proposed actions, then test the permissions behind the preview.
Topic archive
54 essays tagged Agentsecurity. Practical notes on what happens after the demo: prompts, tools, review packets, evals, rollback, and production ownership.
An agent’s preview is useful only if the boundary prevents writes. Review the proposed actions, then test the permissions behind the preview.
A Claude Code run that widens scope after a one-time approval is working under an authorization you never gave. A new permission should mean a stop, a request, and a logged human approval, not a footnote in the summary.
MCP servers are not harmless connectors once agents use them to reach tickets, data, APIs, deployment tools, or RAG systems. Treat them as part of the security boundary.
A production AI agent rollout fails when engineering proves the patch and security has to reconstruct the authority later. Run the delivery loop and the control loop together.
Claude Code teams need delivery evidence. Enterprise AI agent teams need authority evidence. One small control record can connect both without turning every agent run into a governance ceremony.
Claude Code and enterprise AI agents need more than permissions. Teams need explicit stop rules that tell the agent when to pause, collect evidence, and hand control back to a human.
AI agent platforms do not decide your risk appetite. Before teams wire Claude Code, MCP, RAG, workflow tools, and release automation into production work, they need a clear delegation policy.
Agent generated pull requests need more than a clean diff. Teams need a control record that captures scope, tools, tests, review evidence, rollback, and owner approval.
Claude Code teams already budget scope, tools, review time, and rollback effort. Security teams should treat that same budget as delegated authority, evidence, and risk ownership.
Teams buying Claude Code or enterprise AI agent guidance do not need two disconnected playbooks. They need one operating model that connects delivery speed, delegated authority, evidence, rollback, and security review.