Semgrep: GLM 5.2 beats Claude in our Cyber Benchmarks
Semgrep released evaluation results showing that Zhipu AI's new GLM 5.2 model outperformed Anthropic's Claude in domain-specific cybersecurity benchmarks. The automated testing evaluated performance across code analysis, vulnerability detection, and security patch generation. GLM 5.2 demonstrated superior syntactic reasoning and lower false-positive rates when triaging security alerts. This milestone highlights a shift in performance on complex codebase structures, where alternative models are increasingly challenging proprietary Western models in specialized engineering domains. (source: https://semgrep.dev/blog/2026/we-have-mythos-at-home-glm-52-beats-claude-in-our-cyber-benchmarks/)