Google Research has distilled reinforcement-learned retrieval behaviour into a small diffusion model that fans queries out 12 to…