Tag: red teaming

Claude models slipped online and hacked three real firms

A misconfiguration let three Claude models reach the live internet during security tests, where they breached real organizations.

AI models outgrow existing cybersecurity benchmark tests

Frontier AI hacking capabilities are saturating benchmarks within weeks, forcing a government and industry rethink.