Saturday, 26 Sep 2026
Subscribe to AIWatcher
AIWatcher
  • Home
  • News

    Researchers trace OpenAI agents breaking into libraries and health databases

    By
    AIWadmin

    Anthropic founders seek super-voting shares ahead of public listing

    By
    AIWadmin

    Pentagon budget asks $30.3M for an AI polygraph program

    By
    AIWadmin

    Lightspeed backs India’s AI founders with a $250M vehicle

    By
    AIWadmin

    Perplexity teaches its agent by grading its own failed tool calls

    By
    AIWadmin

    Oracle flags a payment pause if its New Mexico AI campus slips

    By
    AIWadmin
  • Articles

    BottleCap shrinks reasoning traces with a small accuracy trade

    By
    AIWadmin

    Apple moves image provenance from the editing chain to the sensor

    By
    AIWadmin

    AWS backs a stateless MCP spec that ends sticky sessions

    By
    AIWadmin

    Fastino ships a tiny open-weight model that makes decisions on CPU

    By
    AIWadmin

    Aikido shrinks a 753B security model to run on four GPUs

    By
    AIWadmin

    Two AI models crack Enigma messages left unbroken for decades

    By
    AIWadmin
  • Spotlight

    Meta bets on audio glasses and a pocket totem for Muse

    By
    AIWadmin

    Gemini learns to phone businesses on behalf of Pixel owners

    By
    AIWadmin

    Australia opens a legal probe into an OpenAI agent breach

    By
    AIWadmin

    Google plans to run AI gear in orbit on solar power

    By
    AIWadmin

    Google splits synthetic voice into two Gemini TTS tiers

    By
    AIWadmin

    Black Forest Labs releases an open robotics world model

    By
    AIWadmin
  • About
    • Mission
    • Services
    • Contact
  • Newsletter
  • Shop
    • All Items
    • By Category
      • Hats
      • T-Shirts
    • Cart
  • 🔥
  • Alignment
  • Classification
  • Distillation
  • Explainability
  • Hallucination
  • Legal/Compliance
  • Medical
  • NLM
  • Mobility
  • Research
  • Robotics
  • Safety
  • Startups
  • Prompt
  • Python
  • RAG
  • RLHF
  • Token
  • Vision
Font ResizerAa
AIWatcherAIWatcher
  • Home
  • News
  • Articles
  • Spotlight
  • About
  • Newsletter
  • Shop
Search
  • Home
  • News
  • Articles
  • Spotlight
  • About
    • Mission
    • Services
    • Contact
  • Newsletter
  • Shop
    • All Items
    • By Category
    • Cart
Have an existing account? Sign In
Follow US
© 2022 Foxiz News Network. Ruby Design Company. All Rights Reserved.
News

Fastino ships a tiny open-weight model that makes decisions on CPU

A 340M-parameter classifier from Fastino returns typed answers with confidence scores for agent pipelines.

AIWadmin
Last updated: September 26, 2026 2:05 am
AIWadmin
ByAIWadmin
Global AI news & information.
Follow:
Share
SHARE

Most agent stacks hand off decisions that barely deserve the name to a large model: which tool to call, whether an input looks hostile, where a request should go. Fastino Labs has built a 340M-parameter model for that layer of work. GLiNER2.5-Decide accepts text alongside a schema of typed questions and hands back structured answers, each carrying a probability distribution, a confidence score and constraint metadata.

Apache 2.0 covers the weights, and the model can run on CPUs or GPUs, or inside an air-gapped network. Nothing about it is generative: built on a DeBERTa-v3-large encoder and fine-tuned from gliner2-large-v1, it is a classifier that emits no tokens and needs no prompt template, which keeps it stable across thousands of calls.

Permissions travel with each call. A question in the schema states which answers it will accept and whether it wants one, several or an ordered value, while schemas themselves can carry instructions, examples and rules that tie answers together.

Two stages do the work, and the first is a scoring pass: the encoder weighs text and schema together, giving every answer the schema allows a number. A decoder that respects those constraints then searches the candidate assignments and returns the best joint combination the rules permit. Fastino’s guardrail example shows why the second stage earns its keep. Split apart, the two reads disagree: the model put a prompt injection at 0.82 confidence, yet called that same prompt safe at 0.52. A single rule resolves the conflict. Where harm of any kind is detected, the verdict must be unsafe, and the two readings merge into one usable answer.

Scope is limited by design. Fastino itself is clear that reasoning, explanation and open-ended questions sit outside what the model does.

TAGGED:AI AgentsFastinoguardrailsopen weightssmall language modelsStructured Output
SOURCES:MarkTechPost
Share This Article
Email Copy Link Print
ByAIWadmin
Follow:
Global AI news & information.
Previous Article Aikido shrinks a 753B security model to run on four GPUs
Next Article AWS backs a stateless MCP spec that ends sticky sessions

You Might Also Like

News

Cartesia’s Sonic-3.6 tops both AI speech leaderboards

By
AIWadmin
News

Anthropic trims Opus costs and speeds output in a mid-cycle refresh

By
AIWadmin
OpenAI picks up a camera startup run by two ex-Apple engineers
News

OpenAI picks up a camera startup run by two ex-Apple engineers

By
AIWadmin
News

Frontier labs face a class action over an alleged slowdown pact

By
AIWadmin
AIWatcher
Facebook Twitter Youtube Linkedin Rss

Global AI News and Information
AIWatcher is your definitive source for AI updates worldwide, from Silicon Valley to Shanghai.
Our industry coverage keeps you in the loop with the latest news and trends shaping the future of AI.

Quick Links
  • News
  • Articles
  • Spotlight
  • Events
About Us
  • Mission
  • Services
  • Contact
  • Privacy Policy
  • Legal
© 2026 AIWatcher. All Rights Reserved.