CybersecurityMachine Learning6 min reading time

Reading Zhipu’s GLM-5.3 results past the headline number

AI News (TechForge)
Read full post
Zhipu's GLM-5.3 model scored highest on the CyberGym benchmark for finding software vulnerabilities, surpassing Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol. However, on more complex exploit reasoning benchmarks, GLM-5.3 lagged behind its American counterparts. The company acknowledges these mixed results and plans to release the model weights publicly.

More in Cybersecurity

Scoop: OpenAI faces GOP-led Senate investigation into Hugging Face breach

Covered by 2 sources

Chinese AI Giants Accused of Sending Millions of User Queries to U.S. Models

The Wall Street Journal
Cybersecurity6 min read

Anthropic Discloses Fourth AI Hacking Incident Involving Claude Opus 4.6

Covered by 2 sources