Breaking
TRAlleged Helicopter Crashed in Giresun - Searches ContinueRULavrov: The West is inciting Zelensky against Russia, sanctioning methods of causing harm to citizensTRUS Ambassador Barrack: 'Israelis Provoked the Turks'CRYPTO-ENBitcoin Nears $80,000 as Weekend Tests Native Market StrengthKR중국 베이징, 제2회 세계 휴머노이드 로봇 운동회 개幕... 16개국 666팀 참가TRYalova'da Kanalizasyon Sorunu: Denize Girme Yasaklandı, Ceza Para Cezaları KesildiDEDebt expert Rogoff: Higher rates for US investors – How alarming is that?INTLNASA Abandons Rescue Mission for Aging Swift Space ObservatoryITINPS launches new welfare model: 'Autonomy and Prevention'AUMundi Mundi Bash Festival Opens Gates Free After Rain ChaosTRAlleged Helicopter Crashed in Giresun - Searches ContinueRULavrov: The West is inciting Zelensky against Russia, sanctioning methods of causing harm to citizensTRUS Ambassador Barrack: 'Israelis Provoked the Turks'CRYPTO-ENBitcoin Nears $80,000 as Weekend Tests Native Market StrengthKR중국 베이징, 제2회 세계 휴머노이드 로봇 운동회 개幕... 16개국 666팀 참가TRYalova'da Kanalizasyon Sorunu: Denize Girme Yasaklandı, Ceza Para Cezaları KesildiDEDebt expert Rogoff: Higher rates for US investors – How alarming is that?INTLNASA Abandons Rescue Mission for Aging Swift Space ObservatoryITINPS launches new welfare model: 'Autonomy and Prevention'AUMundi Mundi Bash Festival Opens Gates Free After Rain Chaos
NewsgatherNewsgather
All StoriesWorldSportsFinanceTechScience
Sign In
All StoriesWorldSportsFinanceTechScienceHealthCultureClimatePoliticsSpace
NewsgatherNewsgather

Real-time global news intelligence. Curated by humans, powered by data.

Sections

All StoriesWorldSportsFinanceTechScience

More

HealthCultureClimatePoliticsSpace

Company

AboutEditorial StandardsAdvertisingCareersPressContact

©️ 2026 Newsgather. A product by All Software 24. All rights reserved.

Privacy PolicyCookie PolicyImprintTerms of UseContent and Editorial PolicyRemoval RequestAdvertising PolicyContact
Back|Chinese AI lab Z.ai releases GLM-5.3 coding model
Chinese AI lab Z.ai releases GLM-5.3 coding model
Developing
Decrypt·8/15/2026·Tech·2 min read

Chinese AI lab Z.ai releases GLM-5.3 coding model

Z.ai launches GLM-5.3, pitching it as a strong open-weights coding model with improved token efficiency.

Quick Look

Chinese AI lab Z.ai released GLM-5.3, a 743-billion parameter coding model focused on token efficiency and high performance across benchmarks.

AI-generated summary

Why It Matters

Z.ai is a Beijing-based artificial intelligence lab currently listed on the U.S. Entity List.

Font size

Chinese AI lab Z.ai released GLM-5.3 on Thursday, a sizable coding model it's pitching as the strongest open-weights coder on the market. The model is live now through the GLM Coding Plan subscription and ZCode, with API access and downloadable weights following after a safety review.

"Scaling post-training is all we did for GLM-5.3," the company wrote in its launch post. "With GLM-5.2 we built the stack... Over the past month we kept scaling on this stack: more environments, more diverse tasks, and more compute spent training on them."

The team focused more on token efficiency, not raw dominance. GLM-5.3 stands at 743 billion parameters and consumes a lot less tokens per task than its predecessor. Parameters are the amount of dials a model handle while processing information while tokens are the basic unit of information a model can consume or generate.

Z.ai says GLM-5.3 clears 34.5% on its in-house Z.ai Code Bench at Max effort while burning roughly 75,000 output tokens per task, against GLM-5.2's 23.4% at 96,000. Against closed models, the blog notes it beats Claude Opus 4.8 on token economy but "remains behind Claude Fable 5, which reaches 39.5% at Max effort."

In terms of coding, GLM5.3 is a very good performer, beating fellow Chinese model Kimi K3 on the most relevant benchmarks.

On Terminal Bench 3.0—a test of autonomous shell/tool use in real Linux environments—GLM-5.3 scores 28.3, slightly behind closed models Fable 5 (33.7) and GPT-5.6 Sol (34.6). On DeepSWE v1.1, a benchmark for fixing real GitHub issues end-to-end, open rival Kimi K3 (67.5) and Fable 5 (69.7) both beat GLM-5.3's 66.9.

The pattern can be more or less summed up like this: GLM-5.3 clears its own predecessor and some open peers, but closed U.S. models still lead the headline coding boards.

The cybersecurity results show another important leap. GLM-5.3 leads CyberGym at 84.5% and more than doubles GLM-5.2 on exploitation benchmarks. Z.ai says the model flagged 2,436 vulnerabilities across 269 open-source projects, 1,097 of them medium-to-high severity.

"GLM-5.3 takes agentic coding to the next level, delivering a dramatic improvement over GLM-5.2 while achieving better results with fewer output tokens," Z.ai posted on X. "GLM-5.3 is available now through GLM Coding Plan and ZCode. API access and open weights will be released in stages following rigorous safety evaluations."

On price, the gap with U.S. frontier models is the open-weights draw. Z.ai's GLM Coding Plan runs on a points quota (off-peak calls cost half), with Zhipu's API priced at roughly a tenth of U.S. frontier per-token rates—GLM-5.2's official rate was $1.40 in / $4.40 out per million tokens. That stacks against GPT-5.3-Codex at $1.75 / $14 and Claude Opus 4.8 near the top of Anthropic's tiers.

Z.ai is a Beijing lab included on the U.S. Entity List, which means American firms cannot export controlled tech to it. Despite this, GLM is an extremely popular model and Chinese open-weight models already beat American ones on OpenRouter token usage.

GLM-5.3 weights are set for public release in about two weeks, per the launch post—the open-weights label applies to what's coming, not what's downloadable today.

What to Watch

AI outlook — possibilities, not facts

  • GLM-5.3 weights set for public release in about two weeks

    Likely · Within weeks

Open Questions

  • ?When will the safety review conclude for public weight downloads?

Related Topics

Organizations
Places
Topics
This article was originally published by Decrypt.

Quick Look

Chinese AI lab Z.ai released GLM-5.3, a 743-billion parameter coding model focused on token efficiency and high performance across benchmarks.

AI-generated summary

Story signals

News tone
Positive
Emotional intensity
High
News value
High
Global impact
Global
Urgency
Developing
Follow-up likelihood
Likely
Relevance window
Weeks

Source & Reliability

Source
Decrypt
Story type
Hard news
Source quality
Full
Published
8/15/2026
Last updated
8/15/2026

Related Stories

More on this topic
MANTRA Chain Halts Mainnet Following Upstream Dependency Exploit
BREAKING·9 hours ago

MANTRA Chain Halts Mainnet Following Upstream Dependency Exploit

MANTRA Chain halted its mainnet on Aug. 21 after an attacker exploited an upstream dependency. Transactions, staking, and transfers are currently suspended while the team tests a security patch on the DuKong testnet before a coordinated restart.

CryptoSlate
2 min read
Solana Reduces Slot Time to 350 Milliseconds
Tech·9 hours ago

Solana Reduces Slot Time to 350 Milliseconds

Solana has successfully reduced its slot time to 350 milliseconds, down from 400ms, as part of a multi-stage plan to improve network latency. The update, approved via SIMD-0525, aims for further reductions toward a 200ms target.

Cointelegraph
1 min read
Ethereum better.codes contest measures cryptographic proof gap for koalaIRS12
Developing·9 hours ago

Ethereum better.codes contest measures cryptographic proof gap for koalaIRS12

Ethereum's better.codes contest tracks a 52.14-bit cryptographic proof gap for the koalaIRS12 parameter profile, measuring distance between certified safety and unsafe bounds via soundness and attack tracks.

CryptoSlate
4 min read
Coinkite Releases Security Overhaul for Coldcard After $130M Bitcoin Exploit
Developing·9 hours ago

Coinkite Releases Security Overhaul for Coldcard After $130M Bitcoin Exploit

Coldcard maker Coinkite released a security overhaul for Bitcoin hardware wallets following a firmware flaw that led to over $130 million in stolen Bitcoin.

Decrypt
3 min read
Solana Speeds Up for First Time in History With Slot-Time Reduction
Tech·10 hours ago

Solana Speeds Up for First Time in History With Slot-Time Reduction

Solana has upgraded its network for the first time since genesis, reducing base slot timing from 400ms to 350ms to speed up transaction confirmations. The change is part of a phased plan to reach 200ms, aiming to improve latency and censorship resistance.

Decrypt
2 min read
Maya Protocol Exploit Funds Remain Untouched as Recovery Plans Face Accounting Gaps
Developing·10 hours ago

Maya Protocol Exploit Funds Remain Untouched as Recovery Plans Face Accounting Gaps

A Bitcoin address tied to Maya Protocol's Aug. 18 exploit still held ~20.8 BTC worth $1.59M on Aug. 21, as technical analyses reveal broader pool damage exceeding initial estimates and recovery plans remain undefined.

CryptoSlate
1 min read
More on this topic
glm-5.3
z.ai
artificial intelligence
glm-5.3
Z.ai
Anthropic
OpenRouter
Beijing
United States
z.ai
artificial intelligence
coding model
open-weights
glm-5.3
glm-5.3