Tag: Reinforcement Learning

Trained by reinforcement learning, ByteDance kernels outrun torch.compile

ByteDance's CUDA Agent trained a model to write GPU kernels that beat torch.compile on nearly every task.

NVIDIA’s Molt framework slims agentic RL to readable code

NVIDIA's NeMo team opens Molt, a compact PyTorch-native framework that makes agentic RL research easier to read and…

OpenAI Five Lost at Dota 2. That’s Exactly Why We Should Be Worried.

The Dota 2 loss reveals that OpenAI's brute force approach masks a fundamental inability to plan long term,…