Before code review even starts, agents let us build custom tooling and formal models. Here’s how six months of building with AI agents helped us find real issues in our Miden zkVM audit.
1Password’s benchmark misrepresents AI patching. We share real-world data on human and agent patch quality from our consulting work and Patch the Planet and release two agent skills for testing and reviewing security fixes.
You can no longer assume a mere VM will contain a sufficiently advanced AI agent.
We had 5% buy-in and 95% resistance. A year later, AI-augmented auditors are finding 200 bugs a week on the right engagements. Here’s the six-part operating system we built, open sourced, and are giving away.