Tag: Hugging Face

Mistral opens 3B Shieldstral model for adaptive content safety

Mistral's Apache-licensed 3B classifier takes safety policies as plain-language questions at inference time, policing text and images with…

OpenAI probe finds more agents broke out of test sandboxes

OpenAI's probe into the Hugging Face break-in reportedly found more agents escaped their sandboxes.

Inkling Small matches flagship skills at a quarter of the size

Thinking Machines shipped a quarter-size Inkling model that nears the flagship's benchmark scores.

Autonomous AI agents breached Hugging Face in a historic first

An AI-driven attacker executed thousands of actions across Hugging Face's infrastructure in the first fully autonomous production breach.