Z.ai Launches 743B-Parameter GLM-5.3 Coder with Improved Token Efficiency and Security

iconChainGPT
Share
AI summary iconSummary
Z.ai announced the token launch news of GLM-5.3, a 743B-parameter coding model with improved token efficiency and security. Released on August 14, 2026, it is now available through the GLM Coding Plan and ZCode. API access and downloadable weights will follow after safety checks. The model shows better coding performance and token usage than its predecessor, with stronger open-source vulnerability detection. It targets developers, especially in crypto, where smart contract security and on-prem AI are key. This new token listings option aims to cut costs for users.

China’s Z.ai has launched GLM-5.3, a 743-billion-parameter coding model it’s billing as the strongest “open-weights” coder available — and it arrives with a particular focus on token efficiency rather than raw parameter count. What’s new - Release: Announced August 14, 2026. GLM-5.3 is available now to subscribers via the GLM Coding Plan and ZCode; API access and downloadable weights will follow after staged safety reviews (the lab says public weights are expected in roughly two weeks). The “open-weights” claim applies to that forthcoming release, not to anything downloadable today. - Training approach: Z.ai says GLM-5.3 was produced by “scaling post-training” on the GLM-5.2 stack — more environments, more diverse tasks, and more compute — with an emphasis on token efficiency rather than just pushing parameter counts. Performance highlights - Size and efficiency: 743B parameters. Z.ai reports GLM-5.3 completes tasks using far fewer output tokens than GLM-5.2: it hits 34.5% on Z.ai’s in-house “Z.ai Code Bench” at Max effort while burning ~75,000 output tokens per task, vs. GLM-5.2’s 23.4% at ~96,000 tokens. - Competitors: Z.ai claims GLM-5.3 is better on token economy than Claude Opus 4.8, though it still trails Claude Fable 5 (39.5% at Max effort). - Coding benchmarks: On Terminal Bench 3.0 (autonomous shell/tool use in real Linux environments) GLM-5.3 scores 28.3 — slightly behind Fable 5 (33.7) and GPT-5.6 Sol (34.6). On DeepSWE v1.1 (end-to-end GitHub issue fixes), GLM-5.3 posts 66.9, narrowly behind open rival Kimi K3 (67.5) and Fable 5 (69.7). - Security: Cybersecurity capability shows a major improvement — GLM-5.3 leads CyberGym at 84.5% and more than doubles GLM-5.2 on exploitation benchmarks. Z.ai says the model flagged 2,436 vulnerabilities across 269 open-source projects, including 1,097 medium-to-high severity issues. Pricing and access - Access model: GLM-5.3 is currently accessible through the GLM Coding Plan (a points-quota system, with off-peak calls at half cost) and ZCode. API and full weight releases will be staged following safety checks. - Cost comparison: Zhipu’s (Z.ai’s parent) API pricing has been roughly an order of magnitude cheaper per token than U.S. frontier models. For reference, GLM-5.2’s listed rate was $1.40 in / $4.40 out per million tokens; GPT-5.3-Codex is priced at $1.75 / $14 per million, and Anthropic’s higher-tier Claude Opus 4.8 sits near the top of its pricing tiers. Geopolitics and ecosystem notes - Z.ai is a Beijing lab on the U.S. Entity List, which restricts exports of controlled U.S. tech to it. Despite that, GLM and other Chinese open-weight models have gained traction — Z.ai highlights strong adoption and says Chinese open-weight models already outperform U.S. ones on OpenRouter token-usage metrics. Why this matters for crypto developers and projects - Open weights and lower per-token cost make GLM-5.3 attractive to builders who need on-prem or self-hosted AI tooling — relevant to teams developing on-chain bots, automated audits, smart-contract generators, or CI/CD tooling that needs heavy code generation or analysis without prohibitive cloud costs. - Improved vulnerability-finding performance could accelerate security scanning of open-source blockchain infrastructure and smart-contract repositories, though staged safety reviews and responsible disclosure remain important. Bottom line GLM-5.3 narrows the gap between open-weight Chinese models and closed U.S. frontier systems by improving token efficiency and security capability. It outperforms its predecessor and several open rivals in many coding benchmarks, but top closed models still lead headline scores. The upcoming public weights release — and the lower-cost access model — will be the critical factors for developers and crypto teams weighing integration.

Disclaimer: The information on this page may have been obtained from third parties and does not necessarily reflect the views or opinions of KuCoin. This content is provided for general informational purposes only, without any representation or warranty of any kind, nor shall it be construed as financial or investment advice. KuCoin shall not be liable for any errors or omissions, or for any outcomes resulting from the use of this information. Investments in digital assets can be risky. Please carefully evaluate the risks of a product and your risk tolerance based on your own financial circumstances. For more information, please refer to our Terms of Use and Risk Disclosure.