Anthropic says automated alignment agents closed more of the safety gap than expert humans on seven of seven…
The University of Toronto number theorist is joining the lab's AI safety team.