Zhipu AI Releases GLM-5.3, Claims It's the Strongest Open-Weights Coding Model
Zhipu AI has released GLM-5.3, a coding-focused model built on the same base as GLM-5.2 with additional post-training. The company claims it's the strongest open-weights coding model available, with gains concentrated in agentic and cybersecurity tasks, though independent benchmarks are not yet published.
Zhipu AI has released GLM-5.3, a new version of its coding-focused language model that the Chinese startup claims is the strongest open-weights coding model currently available. The model shares its base architecture with predecessor GLM-5.2, with all reported improvements coming from extended post-training rather than a new pretraining run.
According to Zhipu, the biggest performance jumps appear in agent-based coding tasks. The company has not published specific benchmark scores, parameter counts, context window size, or pricing for GLM-5.3, so these figures cannot be verified at this time.
Cybersecurity focus
One consistent gap between Chinese frontier models and their US counterparts has been cybersecurity capability. Zhipu says it specifically trained GLM-5.3 with data and simulated environments designed for vulnerability discovery. According to Z.ai, the model "began to reason across multiple stages of exploitation, forming coherent plans for complete exploitation chains" — a claim that has not been independently verified.
Working with security teams in China, Zhipu says the model helped identify 2,436 vulnerabilities across 269 software projects, some of which the company says are up to 40 years old. Zhipu states these findings are documented in a public registry, though the company has not detailed the methodology used to validate each vulnerability or confirm they were previously unknown.
Availability
GLM-5.3 is available now through Zhipu's GLM Coding Plan and is compatible with coding agents including ZCode, Claude Code, and OpenCode. Pricing for the coding plan has not been disclosed in the announcement.
Model weights are set to be released as open source in approximately two weeks, once Zhipu completes internal security reviews. The company has not specified a license type for the eventual open release.
What this means
Zhipu's claim of "strongest open-weights coding model" comes without accompanying benchmark data — no SWE-bench, HumanEval, or agentic coding scores were shared alongside the release. Until Zhipu publishes numbers or third parties run independent evaluations, the claim should be treated as marketing rather than an established result.
The cybersecurity angle is the more concrete development here. If the vulnerability-finding numbers hold up under scrutiny, it would mark a real capability shift for a Chinese open-weights model in a domain where US labs like Anthropic and OpenAI have generally held an edge. The two-week delay before weights ship, explicitly tied to security review, suggests Zhipu itself is treating the exploit-generation capability as something that needs vetting before broad release — a notable admission for a model being marketed on its offensive security skills.
For developers, the practical takeaway is narrower: GLM-5.3 is accessible today only through Zhipu's paid coding plan and select agent integrations, with open weights still pending. Anyone evaluating it against Claude, GPT, or Qwen coding models will need to wait for both the open release and independent benchmark runs to make an informed comparison.
Related Articles
NVIDIA Releases Nemotron 3.5 Lightning 30B-A3B: 3B-Active MoE Model With 1M-Token Context, Quantized for Single-GPU Depl
NVIDIA has published NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4, a 30-billion-parameter Mixture-of-Experts model with only 3B active parameters, a hybrid Mamba-2/MoE/Attention architecture, and support for up to 1 million tokens of context. The NVFP4-quantized checkpoint is designed to run on a single DGX Spark (GB10) or H100 GPU.
Liquid AI Releases LFM2.5-VL-3B, a 3B-Class Vision-Language Model Built for On-Device Deployment
Liquid AI has released LFM2.5-VL-3B, a multimodal upgrade to its LFM2-VL-3B model built for on-device grounding, object detection, and document OCR. The model runs at 228 tokens/sec on an Apple M5 Max and 116 tokens/sec on an AMD Ryzen AI Max+ 395, using under 3.3 GB of memory.
Google DeepMind Ships Gemini 3.7 Flash, Closing Gap With Claude 4.8 and GPT-5.5
Google DeepMind has released Gemini 3.7 Flash, a new entry in its fast-tier model line that reportedly closes a performance gap that opened up under Gemini 3.5 and 3.6 Flash against Anthropic's Claude 4.8+ and OpenAI's GPT-5.5+ series. Full pricing and benchmark details have not yet been disclosed.
Writer Launches Palmyra X6, an Open-Source-Based Model Aimed at Cutting Token Costs 50%
Writer released Palmyra X6, a post-trained variant of Z.ai's open-source GLM-5.2 model, alongside an upgraded agentic harness. The company claims the combination can cut customer token costs by as much as 50% for basic tasks.
Comments
Loading...