Can a Model Actually Find Your Vulnerabilities?
| 39 min read
AI Security July 2026 benchmarks put GPT-5.6 and Kimi K3 at 88.5% recall on rediscovering known CVEs, and found that several cheap runs beat one expensive run. A look at what these numbers mean for your security pipeline, and at the harnesses now competing to structure the work.