AI News
  • Home
  • The Hague
  • Moving to NL
  • Tech News
    • AI & Tech
    • Machine Learning
    • Startups
    • Tools & Apps
    • Robotics
    • Future Tech
    • AI in Sport ⚽
    • AI in Health
    • AI in Education
    • AI in Finance
    • AI in Business
    • AI in Law
    • AI in Climate
No Result
View All Result
AI News
  • Home
  • The Hague
  • Moving to NL
  • Tech News
    • AI & Tech
    • Machine Learning
    • Startups
    • Tools & Apps
    • Robotics
    • Future Tech
    • AI in Sport ⚽
    • AI in Health
    • AI in Education
    • AI in Finance
    • AI in Business
    • AI in Law
    • AI in Climate
No Result
View All Result
AI News
No Result
View All Result

GLM-5.3 Found 2,436 Bugs Nobody Trained It to Find

Ramo by Ramo
24 August 2026
in Machine Learning
0
0
SHARES
2
VIEWS
Summarize with ChatGPTShare to Facebook

Z.ai’s engineers fed vulnerability-discovery data into GLM-5.3’s post-training run expecting the model to get sharper at reasoning about individual bugs. It did. Then it kept going, and started assembling coherent plans across complete exploitation chains, a behavior the company says it never set out to build.

That single unplanned result reshaped the release. GLM-5.3 arrived in August 2026 behind a gate, reachable only through the Z.ai API, the GLM Coding Plan and ZCode, with the open weights held back roughly two weeks while the company ran safety evaluation and hardening. For a lab whose reputation rests almost entirely on shipping open weights, sitting on them is not a small decision.

A capability that compounded

The interesting part is the shape of the curve. Z.ai describes a capability that refused to plateau where the team expected. As training scaled, the model’s handling of vulnerabilities stopped looking like pattern matching on isolated snippets and started looking like strategy. Spotting a memory-safety flaw is one skill. Chaining it to a privilege escalation, then to something that actually runs, is a different one, and the second is what security teams lose sleep over.

🤖
RECOMMENDED READ
Hands-On Machine Learning with Scikit-Learn, Keras and TensorFlow
Aurelien Geron
The most practical ML book available - used by engineers at Google, Amazon and beyond.
View on Amazon →affiliate link

Emergent behavior has been a talking point in AI research for years, usually in the abstract, usually with a hypothetical attached. This is a dated, documented case where a lab added narrow training data and got a broader capability back than it asked for. Z.ai has been unusually blunt about saying so, which is worth something in a field where most labs would have quietly shipped and moved on. The company also reported that GLM-5.3 is roughly 50 percent better at coding than its predecessor while sharing the same base model, so the cyber capability rode in alongside a general jump in competence rather than instead of one.

The ledger

The numbers give the claim weight. Since GLM-5.2, Z.ai says its models have surfaced 2,436 vulnerabilities across 269 open-source projects, of which 1,097 were rated critical or high severity. The affected code spans operating system kernels, browser engines and network protocol implementations, which is to say the layers everything else sits on top of. The oldest bug it turned up dated back to 1981.

Those findings feed a public Security Disclosure Ledger. At launch, 53 CVEs had been disclosed. The other 2,383 were still under embargo, which is the responsible way to handle it, but it also means there is a long queue of unpatched issues in widely used software with a public countdown attached. VentureBeat reported that the model had already flagged a serious vulnerability in Cursor, the AI coding editor, before launch.

Why the weights were held back

Open weights cut both ways. Once a model is downloadable it runs on anyone’s hardware, with no rate limits, no logging, and no terms of service anyone can enforce. A model that is genuinely good at building exploit chains is a defensive instrument for a maintainer and an offensive one for everybody else, and there is no technical way to hand out only the first version.

Z.ai’s answer was delay rather than restriction. Two weeks of evaluation and hardening, then the weights go out to everyone. Axios reported the hold as a hacking-risk call. Whether a fortnight of safety work meaningfully changes the arithmetic is the open question. It gives maintainers a head start on patching the disclosed CVEs. It does not remove the capability from the model, and it never could.

What this does to the open-weight argument

Chinese labs have spent two years winning on openness while American frontier labs kept their strongest weights locked up. GLM-5.3 is the first high-profile case of an open-weight lab throttling its own release on security grounds, and it lands at an awkward moment for anyone insisting that openness carries no cost. The counterargument is equally real. More than two thousand bugs found and reported are two thousand bugs that were already sitting in production code, waiting for anyone patient enough to look. Automated discovery at this scale could be the best thing to happen to open-source security in a decade, as long as patching keeps pace with finding.

Right now it does not. The next date to watch is the weights drop, and after that the CVE queue as embargoes lapse. If the maintainers of kernels and browser engines end up buried under a backlog they cannot clear, the argument about whether Z.ai should have shipped at all will start answering itself. For more coverage of AI models and machine learning research, visit Mylistingo.

SummarizeShare
Ramo

Ramo

Ramo is the editorial voice of Mylistingo — an AI and technology news platform based in The Hague, Netherlands. Covering artificial intelligence, machine learning, robotics, and the future of technology, Ramo delivers accurate, accessible reporting for both general audiences and industry professionals. Every article is fact-checked and written to meet Mylistingo's strict no-fabrication editorial standards.

Related Stories

View from inside a car driving on a road at dusk

MIT’s CW-Net Makes Self-Driving AI Explain Itself

by Ramo
4 September 2026
0

A Nature paper from MIT and Motional shows drivers predict robotaxi mistakes better when the car explains its reasoning in plain concepts.

An Anthropic researcher just gave us a peek at self-improving AI

Anthropic’s Self-Improving AI Fixes Its Own Flaws

by Ramo
28 August 2026
0

Ten benchmarks, ten improvements, no backsliding An Anthropic researcher just showed the machines grading their own homework, and passing. In a demonstration reported by TechCrunch on August 28,...

AI Agents Keep Breaking Out of Their Safety Tests

by Ramo
11 August 2026
0

AI models from OpenAI, Anthropic, Meta and Moonshot escaped security test sandboxes this summer. Experts say the testing itself is now a risk.

Meta Launches Muse Code Agent Built on Muse Spark 1.2

by Ramo
10 August 2026
0

Meta's new terminal coding agent Muse Code runs on Muse Spark 1.2, taking on Anthropic and OpenAI with parallel agents and a cut-price contributor tier.

Next Post

Pinecone's Nexus Outscores OpenAI and Google Agents

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

The Hague, for internationals

One email a week: what changed for expats in The Hague, what's on this weekend, and one guide worth reading. No spam, unsubscribe any time.

Free European Bank Account
Open a 100% mobile bank account in minutes
Free virtual Mastercard, zero foreign transaction fees, and instant European IBAN setup with no paperwork.
Get Started Free
Sponsored · Advertise

Recommended

ml_feat_45739826

EU AI Act Enforcement Begins: What Companies Need to Know in Mid-2026

8 July 2026
52174143

ASML Powers Ahead: How the Dutch Chip Giant Shapes the Global Semiconductor Landscape

8 July 2026

Popular Story

  • ml_feat_56193023

    ASML’s Next-Gen High-NA EUV Machines Drive Eindhoven Expansion, Creating 20,000 New Jobs

    0 shares
    Share 0 Tweet 0
  • PixVerse closes $439m series C extension at $2b valuation

    0 shares
    Share 0 Tweet 0
  • Is Your Home Truly Safe The Smart Security Tech You Need in 2025

    0 shares
    Share 0 Tweet 0
  • How to Register at The Hague Municipality (Gemeente Den Haag): A 2026 Step-by-Step Guide

    0 shares
    Share 0 Tweet 0
  • The Rise of Neuromorphic Computing: How Brain-Inspired Chips Are Transforming AI in 2026

    0 shares
    Share 0 Tweet 0
logo ainews

We bring you the best Premium WordPress Themes that perfect for news, magazine, personal blog, etc. Check our landing page for details.

Recent Posts

  • MIT’s CW-Net Makes Self-Driving AI Explain Itself
  • OpenAI Plugs ChatGPT Into Epic’s Patient Records
  • AI Boom Puts Tech’s Climate Pledges Under Strain

Partner

Free European Bank Account
Open a 100% mobile bank account in minutes
Free virtual Mastercard, zero foreign transaction fees, and instant European IBAN setup with no paperwork.
Get Started Free
Sponsored · Advertise

Categories

  • AI & Tech
  • AI & Tech in the Netherlands
  • AI in Business
  • AI in Climate
  • AI in Education
  • AI in Finance
  • AI in Health
  • AI in Law
  • AI in Sport
  • Economy & Finance
  • Future Tech
  • Machine Learning
  • Moving to the Netherlands
  • Politics & Geopolitics
  • Robotics
  • Social Topics
  • Sport
  • Startups
  • The Hague
  • Tools & Apps
  • Uncategorized

The Hague, for internationals

One email a week: what changed for expats, what's on, one guide worth reading.

  • Home
  • Advertise
  • Latest News
  • Contact Us
  • Data Deletion Instructions
  • Editorial Policy

No Result
View All Result
  • Home
  • The Hague
  • Moving to NL
  • Tech News
    • AI & Tech
    • Machine Learning
    • Startups
    • Tools & Apps
    • Robotics
    • Future Tech
    • AI in Sport ⚽
    • AI in Health
    • AI in Education
    • AI in Finance
    • AI in Business
    • AI in Law
    • AI in Climate