← Back to Model Beat
Models·Jun 28·all news from June 28, 2026

Semgrep Benchmarks GLM-5.2 Against Claude, Finds Higher IDOR F1

Semgrep researchers recently benchmarked the GLM-5.2 model against Anthropic's Claude, finding the former achieved a higher F1 score in identifying insecure direct object reference vulnerabilities. This comparison highlights an increasing trend of specialized, smaller models demonstrating competitive performance in automated security testing compared to larger, general-purpose LLMs.

ModelsGLM-5.2

Covered by 2 sources

Related stories

ModelsChina’s Z.ai claims it can match Mythos on cybersecurityJun 22 · 18 sourcesOpen SourceDatabricks makes Chinese open-source model GLM 5.2 its default coding engine after it matched Opus at lower costJul 9 · 2 sourcesHardwareZhipu AI explores custom ASIC chip as GLM-5.2 usage surges 27x - The InformationJul 7 · 17 sourcesModelsChina’s GLM-5.2 Matches Anthropic’s Mythos Where Cyber Power Matters MostJun 29 · 7 sources