Enterprise DNA

Omni by Enterprise DNA

Enterprise DNA Resources

Latest AI and industry news. Practical AI operating-system thinking for owners, operators, and teams doing real work.

220k+

Data professionals

Omni

AI agents and apps

Audit

Map the manual work

News AI News

Z.ai's GLM-5.3 claims a 50% coding jump and edges frontier labs on a cybersecurity benchmark

Zhipu/Z.ai says GLM-5.3 beats GLM-5.2 by 50% on coding capability from post-training alone, and reports it narrowly leads Anthropic's Mythos and.

Enterprise DNA |
Z.ai's GLM-5.3 claims a 50% coding jump and edges frontier labs on a cybersecurity benchmark

AI Pulse · AI Trends Pulse

The play

Benchmark available models on your actual coding and security tasks, while waiting for weights and independent validation before switching.

Z.ai, the company behind the GLM model family, is claiming a big jump in coding ability with its new GLM-5.3 model. According to the discussion on Hacker News, the company says GLM-5.3 beats its own predecessor, GLM-5.2, by 50% on coding capability, and that gain came purely from post-training work, not from a bigger or different base model. That’s worth noting because it means labs are still finding real gains by refining models after the initial training run, not just by throwing more compute at bigger models.

The other claim getting attention is on security. Z.ai says GLM-5.3 edges out Anthropic’s Mythos and OpenAI’s GPT-5.6 Sol on a benchmark called CyberGym, which tests how well a model handles cybersecurity tasks. The scores are close, 84.5% for GLM-5.3 against 83.8% and 83.6% for the other two. That’s a narrow lead, not a blowout, and these are self-reported numbers from the company, so treat them as a claim rather than an independently confirmed result.

For business owners, the practical point is this. If GLM-5.3 holds up once people can actually test it, it adds another strong, cheaper option to the coding and security tooling market, which keeps pressure on pricing across the board for anyone buying AI capability. Nothing is public yet though. Weights aren’t released, and Z.ai says it needs about two weeks for safety hardening before that happens. So there’s nothing to plug in today. Worth watching once real users get their hands on it and start comparing it against what you’re already using. This is the kind of competitive shift we track for clients inside an AI command centre, so tools that actually move the needle get flagged before the hype settles.

Working With Claude field guide cover

Free Resource

Put what you just read to work

The free 32-page Working With Claude guide: the full ecosystem, Claude Code, and how to roll it out across a business.

No spam. Unsubscribe any time.

Free daily email

Get this every morning.

This brief is one item from today's AI Pulse, the short daily read we run for ourselves on what is actually happening in AI. Subscribe free and it lands in your inbox each morning.

Free daily email

Subscribe to the daily AI Pulse

One short read every morning on what is actually happening in AI. Free.

One email a day. Unsubscribe any time.