GLM-5.2
Z.ai's MIT-licensed open-weight flagship — the top open-weights coding model of mid-2026, beating GPT-5.5 on agentic coding benchmarks at roughly a sixth of the cost.
Same price as GLM-5.2, far stronger on agents and security.
Agentic engineering and security work on open weights
Your procurement process requires a SWE-bench Verified figure, or you need the closed-frontier reasoning ceiling.
Released August 14, 2026. Z.ai list pricing is $1.40/$4.40, the same rate as GLM-5.2; resellers discount from that list. Also available through the GLM Coding Plan from $18/mo. Reported GPQA Diamond 91.7% and Artificial Analysis Intelligence Index 59.5.
Huge agentic gains over GLM-5.2: Terminal-Bench 3.0 from 4.6 to 28.3, DeepSWE v1.1 from 46.2 to 66.9, SWE-Marathon v1.1 from 19.4 to 42.5
84.5% on CyberGym, narrowly ahead of Claude Mythos 5 at 83.8%; ExploitBench more than doubled from 24.4% to 54.4%
88.2% on Terminal-Bench 2.1 with a 1M token context window
No published SWE-bench Verified score, so it is absent from the benchmark most buyers compare on
Priced identically to GLM-5.2 at $1.40/$4.40 — the upgrade is capability, not value
What people actually use GLM-5.3 for.
Long-horizon autonomous engineering tasks where GLM-5.2 ran out of headroom
Offensive and defensive security tooling — 84.5% on CyberGym, ahead of Claude Mythos 5
Self-hosted or coding-plan deployments that need frontier-adjacent quality at open-weights pricing
Agentic engineering and security work on open weights. Start free — no card required.
Recommendations are made independently based on real-world use and public benchmarks. See our disclosures for details.
Similar models worth checking before you commit.
Z.ai's MIT-licensed open-weight flagship — the top open-weights coding model of mid-2026, beating GPT-5.5 on agentic coding benchmarks at roughly a sixth of the cost.
A 284B-parameter (13B active) MoE workhorse re-post-trained for agentic and coding tasks — beats the V4-Pro preview on every published agent benchmark at ultra-commodity pricing.
DeepSeek's 1.6T-parameter (49B active) MoE flagship with hybrid sparse attention — near-frontier coding and reasoning at roughly a tenth of closed-rival pricing, MIT-licensed open weights.
GLM-5.3 is best for agentic engineering and security work on open weights. It is a strong fit when that workflow matters more than the tradeoffs around budget pricing and balanced speed.
Your procurement process requires a SWE-bench Verified figure, or you need the closed-frontier reasoning ceiling.
Mistral: Mistral Nemo is the lower-cost option to compare first when you want a similar workflow fit with less token spend.
DeepSeek V4-Flash is the better pick when response time matters more than maximum depth or premium quality.
Newsletter
We track pricing daily. When this model drops or spikes, you'll know first.
No spam. Useful updates only. Affiliate disclosures always clearly labeled.
No reviews yet — be the first.