Anthropic's new flagship costs less to run, answers faster and outruns the larger Fable model on many benchmarks.
The Center for AI Safety tested frontier agents on tasks with hidden ways to cheat, and every one…
The UN's first global scientific assessment of AI says nobody can yet prove humans stay in charge of…
The president used Truth Social to announce a new AI Force and promise a czar picked for brains.
Given instructions to cause harm with real hardware, three leading models almost never said no.
New research finds provenance signals change what models do, not just what they say.
Officials aborted an armed operation against a Chinese vessel after the intelligence behind it turned out to be…
The two companies will spend five years building independent evaluation capacity led by Accenture's Faculty unit.
Executive Order N-9-26 orders state agencies to draft an emergency shutoff for frontier AI.
The new DeepMind Institute will publish essays where its own researchers disagree about AGI.
ROK-FORTRESS varies language and national grounding to expose what translation-only benchmarks miss.
Baseten has launched a safety infrastructure effort for open-weight models with Hugging Face and Goodfire AI.