ZhipuZ.aiOpenAI360Anthropic

China's Z.ai says GLM-5.3 nears Anthropic's Mythos 5 in cyber-defence tests; scores 84.5% on CyberGym, above Mythos' 83.8%

Chinese AI firm Z.ai announced that its GLM-5.3 model scored 84.5% on the CyberGym cybersecurity test, surpassing Anthropic's Mythos 5 at 83.8%. However, GLM-5.3 lagged in exploit capabilities, scoring 54.4% on ExploitBench, compared to Mythos' 78%.

scmp.com scmp.com+1 source14 August 2026 · 12:21 UTC
CuriousCats Full Story

Chinese AI startup Z.ai has unveiled its GLM-5.3 model, claiming it achieved an 84.5% success rate on the CyberGym test, surpassing Anthropic's Mythos 5 at 83.8%. This benchmark evaluates a model's ability to identify and validate security flaws in source code.123

Zhipu, the company behind Z.ai, stated that GLM-5.3 was tested against real-world codebases, identifying 2,436 vulnerabilities across 269 projects, with 1,097 rated medium to high severity. However, the model fell short in exploit capabilities, scoring 54.4% on the ExploitBench test, compared to Mythos' 78% and OpenAI's GPT-5.6 Sol at 76.5%.47

In a timed test, GLM-5.3 completed 105 attack-development tasks in two hours and 130 in six hours, while Mythos 5 completed 181 and 247 tasks respectively. Z.ai plans to release GLM-5.3 publicly in two weeks, with enhanced security measures and a "trusted access" program for sensitive functions.5

Z.ai's launch challenges the restricted access model of Mythos, advocating for open-source tools in cybersecurity. The company aims to support smaller security teams and developers by initiating an "Open Source Shield" initiative to audit open-source projects and provide model access for defensive work.

Despite the advancements, critics warn that enforcing safeguards becomes challenging once models are publicly available, raising concerns about potential misuse.

Key Insight
“GLM-5.3 trails on ExploitBench, scoring 54.4% versus Mythos 5's 78%, and completed 105 attack-development tasks in two hours versus Mythos' 181. Z.ai plans a public release in about two weeks, with sensitive functions under a 'trusted access' programme.”
CuriousCats studied:
1
scmp.comscmp.com
“Chinese artificial intelligence firm Zhipu, also known as Z.ai, has unveiled its flagship GLM-5.3 model, saying it beat Anthropic’s frontier Mythos 5 model in a key cybersecurity test, as China races to counter Western advances in AI defence.”
scmp.com →
2
CNACNA
“BEIJING: Chinese AI startup Z.ai said on Friday (Aug 14) its open-source GLM-5.3 model had neared Anthropic's restricted Mythos 5 in identifying software vulnerabilities, bolstering the credentials of a Chinese AI challenger gaining traction among Western developers.”
CNA →
Ask CuriousCats
Who developed GLM-5.3?
What are the scores for CyberGym tests?
Why is ExploitBench performance important?
Are there other AI models competing in cyber-defence?
How does Z.ai plan to restrict sensitive functions?
Get your CIA-level briefing,
in real time.
CuriousCats monitors the internet every minute for you and brings you the most personalized brief of videos, social media posts, news and more.
Download the App
If you liked this, you’ll love your CuriousCats brief.
News, videos, opinions and more — without the noise.
Get CuriousCats