Tag: self-improving AI

Jacob Coxon leaves Anthropic over self-improving AI fears

A researcher who trained models at OpenAI and Anthropic has quit, calling the race to self-improving AI a…

Claude-built researchers beat humans at fixing model flaws

Anthropic says automated alignment agents closed more of the safety gap than expert humans on seven of seven…