Zhipu AI Releases GLM-5.3, Claims It's the Strongest Open-Weights Coding Model
Zhipu AI has released GLM-5.3, a coding-focused model built on the same base as GLM-5.2 with additional post-training. The company claims it's the strongest open-weights coding model available, with gains concentrated in agentic and cybersecurity tasks, though independent benchmarks are not yet published.
GLM-5.3 — Quick Specs
Zhipu AI has released GLM-5.3, a new version of its coding-focused language model that the Chinese startup claims is the strongest open-weights coding model currently available. The model shares its base architecture with predecessor GLM-5.2, with all reported improvements coming from extended post-training rather than a new pretraining run.
According to Zhipu, the biggest performance jumps appear in agent-based coding tasks. The company has not published specific benchmark scores, parameter counts, context window size, or pricing for GLM-5.3, so these figures cannot be verified at this time.
Cybersecurity focus
One consistent gap between Chinese frontier models and their US counterparts has been cybersecurity capability. Zhipu says it specifically trained GLM-5.3 with data and simulated environments designed for vulnerability discovery. According to Z.ai, the model "began to reason across multiple stages of exploitation, forming coherent plans for complete exploitation chains" — a claim that has not been independently verified.
Working with security teams in China, Zhipu says the model helped identify 2,436 vulnerabilities across 269 software projects, some of which the company says are up to 40 years old. Zhipu states these findings are documented in a public registry, though the company has not detailed the methodology used to validate each vulnerability or confirm they were previously unknown.
Availability
GLM-5.3 is available now through Zhipu's GLM Coding Plan and is compatible with coding agents including ZCode, Claude Code, and OpenCode. Pricing for the coding plan has not been disclosed in the announcement.
Model weights are set to be released as open source in approximately two weeks, once Zhipu completes internal security reviews. The company has not specified a license type for the eventual open release.
What this means
Zhipu's claim of "strongest open-weights coding model" comes without accompanying benchmark data — no SWE-bench, HumanEval, or agentic coding scores were shared alongside the release. Until Zhipu publishes numbers or third parties run independent evaluations, the claim should be treated as marketing rather than an established result.
The cybersecurity angle is the more concrete development here. If the vulnerability-finding numbers hold up under scrutiny, it would mark a real capability shift for a Chinese open-weights model in a domain where US labs like Anthropic and OpenAI have generally held an edge. The two-week delay before weights ship, explicitly tied to security review, suggests Zhipu itself is treating the exploit-generation capability as something that needs vetting before broad release — a notable admission for a model being marketed on its offensive security skills.
For developers, the practical takeaway is narrower: GLM-5.3 is accessible today only through Zhipu's paid coding plan and select agent integrations, with open weights still pending. Anyone evaluating it against Claude, GPT, or Qwen coding models will need to wait for both the open release and independent benchmark runs to make an informed comparison.
Related Articles
Z.ai Releases GLM-5.3-Prime, a High-Throughput Variant of GLM-5.3 with 1M-Token Context
Z.ai has released GLM-5.3-Prime, a high-speed variant of its GLM-5.3 model that delivers 1.5-2x the output throughput through inference acceleration while retaining the full 1M-token context window. The model is priced at $2.80 per 1M input tokens and $8.80 per 1M output tokens, targeting coding and long-horizon agentic workloads.
H Company Releases Holo4 Agentic Models, Scoring 61.7% on OSWorld 2.0 at a Fraction of Frontier Cost
H Company has released Holo4, a series of agentic models in 27B dense and 35B-A3B Mixture of Experts sizes that operate across GUIs, code, MCP, and APIs using a single interface. The models score 61.7% (27B) and 30.9% (35B-A3B) on OSWorld 2.0, trailing closed frontier models like Opus 5.5 (81.8%) but at far lower cost.
Nvidia Releases Nemotron 3 Diarization, a Free 100M-Parameter Model That Tracks 8 Speakers in Real Time
Nvidia released Nemotron 3 Diarization, a free 100-million-parameter model that identifies who is speaking in real time across up to eight participants. It leads the VoiceArena Diarization Benchmark v1 with a 14.7% error rate, cutting errors by 41% versus its predecessor.
Apple Releases LensVLM-9B, a 9B Vision-Language Model That Selectively Decompresses Text Images
Apple has released LensVLM-9B, a 9-billion-parameter vision-language model fine-tuned from Qwen3.5-9B-Base that processes documents as compressed images, selectively expanding only relevant pages to full resolution. The model supports 5x, 10x, and 15x compression ratios and is available under Apple's Machine Learning Research Model License.
Comments
Loading...