Tag: GLM 5.2

Perplexity teaches its agent by grading its own failed tool calls

A new post-training method trains Perplexity's computer agent on real sessions, including the ones that went wrong.