AMD's fully open mixture-of-experts model leads open rivals while activating just 2.8B parameters.
AMD ships Instella-MoE-16B-A3B, an open mixture-of-experts model trained on Instinct GPUs with 2.8B active parameters.