A Semgrep Cybersecurity Benchmark Study
Semgrep Research Team
2026
Testing AI models on cybersecurity tasks
Overall benchmark score across cybersecurity tasks
What the results reveal
GLM 5.2 demonstrates that specialized training can outperform general-purpose models in cybersecurity benchmarks — opening new possibilities for security automation