• Space Exploration Technologies (Dinari Tokenized Stock)Space Exploration Technologies (Dinari Tokenized Stock)(SPCX)$139.732.60%
  • bitcoinBitcoin(BTC)$62,712.00-1.50%
  • ethereumEthereum(ETH)$1,871.19-1.10%
  • tetherTether(USDT)$1.000.00%
  • binancecoinBNB(BNB)$607.20-0.90%
  • usd-coinUSDC(USDC)$1.000.00%
  • rippleXRP(XRP)$1.00-0.70%
  • solanaSolana(SOL)$75.40-0.80%
  • tronTRON(TRX)$0.333452-1.10%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.01-3.20%
  • HyperliquidHyperliquid(HYPE)$56.39-1.10%
  • dogecoinDogecoin(DOGE)$0.069432-1.00%
  • USDSUSDS(USDS)$1.000.00%
  • RainRain(RAIN)$0.0128875.20%
  • leo-tokenLEO Token(LEO)$9.22-1.80%
  • zcashZcash(ZEC)$486.78-1.50%
  • moneroMonero(XMR)$396.34-1.20%
  • cardanoCardano(ADA)$0.181714-1.90%
  • chainlinkChainlink(LINK)$8.780.20%
  • whitebitWhiteBIT Coin(WBT)$54.39-1.40%
  • stellarStellar(XLM)$0.158411-1.60%
  • daiDai(DAI)$1.000.00%
  • bitcoin-cashBitcoin Cash(BCH)$204.84-4.60%
  • USD1USD1(USD1)$1.000.00%
  • Ethena USDeEthena USDe(USDE)$1.000.00%
  • CantonCanton(CC)$0.095686-3.20%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.32-1.10%
  • Global DollarGlobal Dollar(USDG)$1.000.00%
  • litecoinLitecoin(LTC)$44.60-0.60%
  • Circle USYCCircle USYC(USYC)$1.130.00%
  • hedera-hashgraphHedera(HBAR)$0.065288-1.70%
  • paypal-usdPayPal USD(PYUSD)$1.000.00%
  • suiSui(SUI)$0.68-1.40%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • avalanche-2Avalanche(AVAX)$6.36-2.10%
  • tether-goldTether Gold(XAUT)$4,336.57-0.40%
  • shiba-inuShiba Inu(SHIB)$0.000004-0.40%
  • crypto-com-chainCronos(CRO)$0.0488114.60%
  • uniswapUniswap(UNI)$3.43-3.90%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.140.10%
  • okbOKB(OKB)$101.20-1.00%
  • nearNEAR Protocol(NEAR)$1.60-4.40%
  • BittensorBittensor(TAO)$199.080.00%
  • pax-goldPAX Gold(PAXG)$4,354.34-0.30%
  • World Liberty FinancialWorld Liberty Financial(WLFI)$0.0540472.00%
  • HTX DAOHTX DAO(HTX)$0.000002-0.30%
  • AsterAster(ASTER)$0.60-0.50%
  • OndoOndo(ONDO)$0.328821-1.10%
  • Ripple USDRipple USD(RLUSD)$1.000.00%
  • usddUSDD(USDD)$1.000.00%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Z.ai Ships GLM-5.3 Without Retraining the Base Model: Better at Complex Coding and Long-Horizon Tasks

August 14, 2026
in AI & Technology
Reading Time: 17 mins read
A A
Z.ai Ships GLM-5.3 Without Retraining the Base Model: Better at Complex Coding and Long-Horizon Tasks
ShareShareShareShareShare

Z.ai just released GLM-5.3. GLM-5.3 runs on the same 743B base model as GLM-5.2. Every reported gain comes from scaled post-training: more task environments, more environment types, longer training. The results land in two places. Coding jumps most on the longest-horizon benchmarks, with Terminal-Bench 3.0 moving from 4.6 to 28.3. Cybersecurity moved further than Z.ai says it expected, with CyberGym reaching 84.5%. Weights are not public yet.

Is It Deployable?

Partially, GLM-5.3 is live through the Z.ai API, the GLM Coding Plan, and ZCode. Weights are not out. Z.ai says it will publish them roughly two weeks after launch, once safety evaluation and hardening finish.

YOU MAY ALSO LIKE

Trump Slaps A 100 Percent Tariff On Heavy And ‘Sensitive’ Drones

Z.ai Launches GLM-5.3 With Frontier Coding and a Cyber Capability That Outgrew Its Training – Unite.AI

  • Which companies can move now: Startups and mid-market engineering orgs can adopt it today via the Coding Plan or API. Enterprises with data-residency or vendor-review rules should wait for weights. Security vendors and MSSPs get the most signal, and the most policy exposure.
  • Industries: Developer tooling, cloud infrastructure, application security, fintech and e-commerce engineering, and vendors shipping kernels, browser engines, or network stacks.
  • Applications: Repository-scale refactors, long-horizon CLI agents, CI failure triage, white-box vulnerability discovery, crash triage, and secure code review.

Coding Results

Terminal-Bench 3.0 moves from 4.6 to 28.3 against GLM-5.2. DeepSWE v1.1 moves from 46.2 to 66.9. Agents’ Last Exam (CLI) moves from 23.8 to 28.5. On GDPval-AA v2, which spans 44 occupations, GLM-5.3 scores 1,769.

On Z.ai Code Bench, an internal evaluation, the company reports a 50% improvement over GLM-5.2. It reports 31.4% at roughly 50,000 output tokens per task. Claude Opus 4.8 scores 29.5% at 120,000 tokens. Claude Fable 5 still leads at 39.5% at maximum effort. Z.ai argues a private benchmark reduces contamination risk.

On public suites, GLM-5.3 trails GPT-5.6 Sol and Fable 5 on several harder coding evaluations. All figures are vendor-reported, with harness, context length, and sampling settings documented in the announcement.

The Cybersecurity Result

Z.ai flags this one as unplanned. It added vulnerability-discovery data expecting better single-bug reasoning. Instead, capability kept compounding as training scaled. The model began forming coherent plans across complete exploitation chains.

CyberGym, which tests discovery and validation from white-box source, moves from 77.2% to 84.5%. That edges past Mythos 5 at 83.8% and GPT-5.6 Sol at 83.6%. ExploitBench, which requires root-cause reasoning and a working exploit, moves from 24.4% to 54.4%. Mythos 5 sits at 78.0%. On ExploitGym, GLM-5.3 completes 105 tasks in two hours and 130 in six. GLM-5.2 completes 29 and 39. Mythos 5 completes 181 and 247.

The pattern is consistent. The deeper into the exploitation chain a benchmark sits, the larger the gain over GLM-5.2. The gap to closed frontier models also widens.

Interactive Explainer


Key Takeaways

  • GLM-5.3 reuses the GLM-5.2 base model; all gains come from post-training scaling.
  • Terminal-Bench 3.0 moves from 4.6 to 28.3; DeepSWE v1.1 from 46.2 to 66.9.
  • CyberGym hits 84.5%, ahead of Mythos 5 (83.8%) and GPT-5.6 Sol (83.6%).
  • ExploitBench more than doubles to 54.4%, but trails Mythos 5 at 78.0%.
  • Weights ship in about two weeks, after safety evaluation and hardening.

Check out the Z.ai GLM-5.3 technical blog, Zai_org announcement, Z.ai Security Disclosure Ledger and zai-org/GLM-5 on GitHub. Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter. Wait! are you on telegram? now you can join us on telegram as well.

Need to partner with us for promoting your GitHub Repo OR Hugging Face Page OR Product Release OR Webinar etc.? Connect with us


Asif Razzaq is the CEO of Marktechpost Media Inc.. As a visionary entrepreneur and engineer, Asif is committed to harnessing the potential of Artificial Intelligence for social good. His most recent endeavor is the launch of an Artificial Intelligence Media Platform, Marktechpost, which stands out for its in-depth coverage of machine learning and deep learning news that is both technically sound and easily understandable by a wide audience. The platform boasts of over 2 million monthly views, illustrating its popularity among audiences.

Credit: Source link

ShareTweetSendSharePin

Related Posts

Trump Slaps A 100 Percent Tariff On Heavy And ‘Sensitive’ Drones
AI & Technology

Trump Slaps A 100 Percent Tariff On Heavy And ‘Sensitive’ Drones

August 14, 2026
Z.ai Launches GLM-5.3 With Frontier Coding and a Cyber Capability That Outgrew Its Training – Unite.AI
AI & Technology

Z.ai Launches GLM-5.3 With Frontier Coding and a Cyber Capability That Outgrew Its Training – Unite.AI

August 14, 2026
Create a Reasoning-Focused LLM: A Practical Guide to Streaming, Curating, and Fine-Tuning the SupraLabs Reasoning Corpus
AI & Technology

Create a Reasoning-Focused LLM: A Practical Guide to Streaming, Curating, and Fine-Tuning the SupraLabs Reasoning Corpus

August 14, 2026
Pony.ai and Uber Expand Partnership to 2,000+ Robotaxis Across Europe – Unite.AI
AI & Technology

Pony.ai and Uber Expand Partnership to 2,000+ Robotaxis Across Europe – Unite.AI

August 14, 2026
Next Post
L.A. schools superintendent resigns amid FBI investigation

L.A. schools superintendent resigns amid FBI investigation

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Full Episode: TODAY Show – July 2

Full Episode: TODAY Show – July 2

August 7, 2026
Supreme Court rules President Trump cannot fire Fed member Lisa Cook

Supreme Court rules President Trump cannot fire Fed member Lisa Cook

August 10, 2026
South Park Commons Raises Ambitions for the AI Era

South Park Commons Raises Ambitions for the AI Era

August 12, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!