Monday, 17 Aug 2026
Subscribe to AIWatcher
AIWatcher
  • Home
  • News

    Needle 2 squeezes tool calling into a 14MB model binary

    By
    AIWadmin

    Dyna-2 robot model learns from 170 years of human video

    By
    AIWadmin

    Kog bets software unlocks 30x faster inference on stock GPUs

    By
    AIWadmin

    Trust deficit, not dire warnings, drives AI backlash, Amodei says

    By
    AIWadmin

    Gas price tripling forecast clouds hyperscaler power plans

    By
    AIWadmin

    DeepSeek hikes V4 prices as Flash flubs complex agent jobs

    By
    AIWadmin
  • Articles

    New $300M fund backs AI for hospitality and venues

    By
    AIWadmin

    Ransomware operators put Claude Code inside live hacks

    By
    AIWadmin

    Claude Code bills for blank thinking blocks users never see

    By
    AIWadmin

    Liquid AI’s 3B vision model brings screen agents to laptops

    By
    AIWadmin

    Okta trims agent token bills by hiding unneeded MCP tools

    By
    AIWadmin

    Samsung trains wearable health models that learn from biosignals

    By
    AIWadmin
  • Spotlight

    Suno Studio 2.0 lets keyboards feed AI tracks with real playing

    By
    AIWadmin

    Google lets creators drop visible watermarks from Gemini output

    By
    AIWadmin

    Anthropic keeps Model 2 in the vault as misalignment risk ticks up

    By
    AIWadmin

    OpenAI crosses $40B run rate as enterprise outgrows consumer sales

    By
    AIWadmin

    Databricks tops $7B run rate as $5B round values it at $190B

    By
    AIWadmin

    SpaceX closes $60B Cursor deal to pair code agents with GPU fleet

    By
    AIWadmin
  • About
    • Mission
    • Services
    • Contact
  • Newsletter
  • Shop
    • All Items
    • By Category
      • Hats
      • T-Shirts
    • Cart
  • 🔥
  • Alignment
  • Classification
  • Distillation
  • Explainability
  • Hallucination
  • Legal/Compliance
  • Medical
  • NLM
  • Mobility
  • Research
  • Robotics
  • Safety
  • Startups
  • Prompt
  • Python
  • RAG
  • RLHF
  • Token
  • Vision
Font ResizerAa
AIWatcherAIWatcher
  • Home
  • News
  • Articles
  • Spotlight
  • About
  • Newsletter
  • Shop
Search
  • Home
  • News
  • Articles
  • Spotlight
  • About
    • Mission
    • Services
    • Contact
  • Newsletter
  • Shop
    • All Items
    • By Category
    • Cart
Have an existing account? Sign In
Follow US
© 2022 Foxiz News Network. Ruby Design Company. All Rights Reserved.
News

GLM-5.3 ships from Z.ai with an unplanned cyber edge

Z.ai's GLM-5.3 skips retraining and scales up post-training, with an unexpectedly strong cyber capability.

AIWadmin
Last updated: August 14, 2026 10:12 pm
AIWadmin
ByAIWadmin
Global AI news & information.
Follow:
Share
SHARE

An unintended cybersecurity capability is the headline of Z.ai’s latest model update. GLM-5.3 keeps the base architecture of GLM-5.2 and concentrates every improvement in post-training, yet vulnerability-handling skill reportedly grew faster than the team expected as compute scaled.

The release is live through Z.ai’s API and GLM Coding Plan, with open weights promised in roughly two weeks once safety checks wrap. The company’s comparison table pits the model against DeepSeek-V4 Pro, Moonshot’s Kimi K3, and OpenAI’s GPT-5.6 Sol across coding, cyber, and agentic benchmarks.

Instead of changing the architecture, Z.ai scaled the training environments. The stack from GLM-5.2, including the IndexShare long-context technique and the SAO reinforcement learning method for long-horizon tasks, was applied to more diverse work-like settings. One example places the model in an ML infrastructure engineer’s environment with compute clusters, documentation, and codebases, where it must find bottlenecks and deliver measurable speedups. Some tasks represent days of work for an experienced engineer.

Benchmark gains concentrate on the longest-horizon tests. Terminal-Bench 3.0 scores jump from 4.6 to 28.3 over GLM-5.2, and DeepSWE v1.1 moves from 46.2 to 66.9. On Z.ai’s own Code Bench, the model shows a 50% improvement over its predecessor and outscores Claude Opus 4.8 at comparable effort while consuming fewer output tokens, though Claude Fable 5 and GPT-5.6 Sol still lead on several public evaluations.

The company introduced vulnerability discovery data into post-training expecting gains in reasoning about individual flaws. Instead, capability grew faster than anticipated, a result Z.ai flagged as surprising in its announcement.

TAGGED:AI benchmarkscoding modelsCybersecurityGLM-5.3open weightspost-trainingZ.ai
SOURCES:Unite AIMarkTechPost
Share This Article
Email Copy Link Print
ByAIWadmin
Follow:
Global AI news & information.
Previous Article Apple weighs pay-per-use deals to feed Siri the news
Next Article Pony.ai and Uber aim 2,000 robotaxis at European streets

You Might Also Like

EventsNews

OpenAI Orders GPT-5.5 to Never Mention Goblins. What Are They Hiding?

By
AIWadmin
News

Beijing Reinvents Itself as China’s “AI First City” with Unmatched Innovation and Global Ambition

By
Zoe Chang
News

Twitch gives streamers control over Amazon AI training data

By
AIWadmin
News

The Xteink X3: A Minimalist E-Ink Hack For Your Doomscrolling Addiction, But It Won’t Save You

By
AIWadmin
AIWatcher
Facebook Twitter Youtube Linkedin Rss

Global AI News and Information
AIWatcher is your definitive source for AI updates worldwide, from Silicon Valley to Shanghai.
Our industry coverage keeps you in the loop with the latest news and trends shaping the future of AI.

Quick Links
  • News
  • Articles
  • Spotlight
  • Events
About Us
  • Mission
  • Services
  • Contact
  • Privacy Policy
  • Legal
© 2026 AIWatcher. All Rights Reserved.