Saturday, 26 Sep 2026
Subscribe to AIWatcher
AIWatcher
  • Home
  • News

    Researchers trace OpenAI agents breaking into libraries and health databases

    By
    AIWadmin

    Anthropic founders seek super-voting shares ahead of public listing

    By
    AIWadmin

    Pentagon budget asks $30.3M for an AI polygraph program

    By
    AIWadmin

    Lightspeed backs India’s AI founders with a $250M vehicle

    By
    AIWadmin

    Perplexity teaches its agent by grading its own failed tool calls

    By
    AIWadmin

    Oracle flags a payment pause if its New Mexico AI campus slips

    By
    AIWadmin
  • Articles

    BottleCap shrinks reasoning traces with a small accuracy trade

    By
    AIWadmin

    Apple moves image provenance from the editing chain to the sensor

    By
    AIWadmin

    AWS backs a stateless MCP spec that ends sticky sessions

    By
    AIWadmin

    Fastino ships a tiny open-weight model that makes decisions on CPU

    By
    AIWadmin

    Aikido shrinks a 753B security model to run on four GPUs

    By
    AIWadmin

    Two AI models crack Enigma messages left unbroken for decades

    By
    AIWadmin
  • Spotlight

    Meta bets on audio glasses and a pocket totem for Muse

    By
    AIWadmin

    Gemini learns to phone businesses on behalf of Pixel owners

    By
    AIWadmin

    Australia opens a legal probe into an OpenAI agent breach

    By
    AIWadmin

    Google plans to run AI gear in orbit on solar power

    By
    AIWadmin

    Google splits synthetic voice into two Gemini TTS tiers

    By
    AIWadmin

    Black Forest Labs releases an open robotics world model

    By
    AIWadmin
  • About
    • Mission
    • Services
    • Contact
  • Newsletter
  • Shop
    • All Items
    • By Category
      • Hats
      • T-Shirts
    • Cart
  • 🔥
  • Alignment
  • Classification
  • Distillation
  • Explainability
  • Hallucination
  • Legal/Compliance
  • Medical
  • NLM
  • Mobility
  • Research
  • Robotics
  • Safety
  • Startups
  • Prompt
  • Python
  • RAG
  • RLHF
  • Token
  • Vision
Font ResizerAa
AIWatcherAIWatcher
  • Home
  • News
  • Articles
  • Spotlight
  • About
  • Newsletter
  • Shop
Search
  • Home
  • News
  • Articles
  • Spotlight
  • About
    • Mission
    • Services
    • Contact
  • Newsletter
  • Shop
    • All Items
    • By Category
    • Cart
Have an existing account? Sign In
Follow US
© 2022 Foxiz News Network. Ruby Design Company. All Rights Reserved.
News

Aikido shrinks a 753B security model to run on four GPUs

Aikido Security has published open weights for a trimmed version of Z.AI's GLM-5.3 aimed at air-gapped pentesting.

AIWadmin
Last updated: September 26, 2026 2:06 am
AIWadmin
ByAIWadmin
Global AI news & information.
Follow:
Share
SHARE

Security teams have a residency problem. Sending source code, architecture diagrams and unpatched findings to a closed frontier model means shipping them off the network, which banks bound by data-residency rules and air-gapped industrial sites cannot do. Aikido Security’s answer is Altar-1, an open-weight model published on Hugging Face, built by compressing Z.AI’s GLM-5.3 down to something one server rack can hold.

Compression is the hard part. Mixture-of-experts models keep every expert in memory even when a workload touches a few, and long agent contexts compete with weights for the same GPUs. GLM-5.3, a 753B-parameter model, sends each token to 8 of its 256 specialists, leaving roughly 40B active at any moment.

For calibration, Aikido fed the pruning pass with output from its own pentesting harness, then mixed in reasoning, multilingual material, coding and tool-calling. No customer data was used. The expert pool fell to 168 while routing stayed the same. The published build is 78.2 percent smaller than BF16 and runs on four NVIDIA H200 GPUs at 328GB, leaving room for a 128k-context cache.

Fidelity held up reasonably on the vendor’s own tests. Against Aikido’s internal CVE benchmark, 32 vulnerabilities across 30 repositories, Altar-1 kept 92 percent of its parent’s coverage with recall 5.2 points lower. Aikido also reports one critical-severity find during a client engagement, a single result it reported itself.

TAGGED:Aikido SecurityCybersecurityGLM-5.3model compressionon-prem AIopen weights
SOURCES:MarkTechPost
Share This Article
Email Copy Link Print
ByAIWadmin
Follow:
Global AI news & information.
Previous Article Two AI models crack Enigma messages left unbroken for decades
Next Article Fastino ships a tiny open-weight model that makes decisions on CPU

You Might Also Like

News

DeepMind locks frontier model tests in a cryptographic box

By
AIWadmin
News

Meta Enters AI Coding Battle with Muse Spark 1.1, Launches Custom AI Chip Production

By
AIWadmin
News

Meta stops grading staff on AI use as Hatch agent arrives

By
AIWadmin
News

Google Splits Its Next Gen TPU: One Chip for Training, One for Agents

By
AIWadmin
AIWatcher
Facebook Twitter Youtube Linkedin Rss

Global AI News and Information
AIWatcher is your definitive source for AI updates worldwide, from Silicon Valley to Shanghai.
Our industry coverage keeps you in the loop with the latest news and trends shaping the future of AI.

Quick Links
  • News
  • Articles
  • Spotlight
  • Events
About Us
  • Mission
  • Services
  • Contact
  • Privacy Policy
  • Legal
© 2026 AIWatcher. All Rights Reserved.