Critique: 1Password's AI Patching Benchmark
A critique of 1Password's August 6, 2026 report says its 26% "clean fix" rate is misleading: the sample focused on difficult bugs, 22% of trials instructed agents to apply wrong fixes, 36% forbade building or testing, and models used different reasoning settings. The authors also released agent skills for post-patch validation and review walkthroughs.