A researcher who trained models at OpenAI and Anthropic has quit, calling the race to self-improving AI a…
Anthropic says automated alignment agents closed more of the safety gap than expert humans on seven of seven…