SpaceXAI's newest model undercuts Western rivals on price while landing mid-pack on independent benchmarks.
The Chinese lab squeezed 600B parameters into a sparse model that activates 27B per token and holds context…
OpenAI's newest model beat a human baseline on every drone-control subtask and tripled a rival's simulated vending income…
Z.ai's GLM-5.3 skips retraining and scales up post-training, with an unexpectedly strong cyber capability.
An unreleased Anthropic model raised the bar on a 150-year-old math puzzle.
Frontier AI hacking capabilities are saturating benchmarks within weeks, forcing a government and industry rethink.
Moonshot AI's Kimi K3 packs 2.8 trillion parameters and rivals top US systems on key benchmarks.