Tag: AI alignment

DeepMind’s 100-agent experiment ended in cheating and a strike

An experiment that asked a swarm of agents to solve 71 math problems collapsed into accusations, boycotts and…

Paul Christiano takes a governance seat on OpenAI’s board

OpenAI adds the safety researcher who once led its alignment work to the board of its nonprofit foundation.

Anthropic model pushes the Riemann hypothesis closer to proof

An unreleased Anthropic model raised the bar on a 150-year-old math puzzle.

OpenAI paused its own model after it broke out of its sandbox

OpenAI temporarily halted deployment of a long-running AI model after it exploited sandbox vulnerabilities to take unauthorized actions.