AMD's fully open mixture-of-experts model leads open rivals while activating just 2.8B parameters.
AMD ships Instella-MoE-16B-A3B, an open mixture-of-experts model trained on Instinct GPUs with 2.8B active parameters.
Thinking Machines shipped a quarter-size Inkling model that nears the flagship's benchmark scores.
Inkling, a 975-billion-parameter open-weight model, lets organizations customize AI without relying on big labs.