Tag: Mixture of Experts

Jina AI fits a document parser into a 3.4B open model

The new jina-ocr-v1 turns PDFs, scans and tables into clean Markdown on a single low-cost GPU.

Cohere bets on translation with a 218B open-weight release

North Small Translate arrives with 218B parameters, aggressive per-task pricing and a licence that stops short of commercial…

DeepSeek graduates V4 Pro to production after a four-month preview

DeepSeek's flagship V4 Pro reached general availability as build 0813, ending a four-month preview with unchanged pricing.

AMD’s open 16B MoE model tops fully open rivals on benchmarks

AMD's fully open mixture-of-experts model leads open rivals while activating just 2.8B parameters.

AMD opens Instella MoE model trained on its Instinct GPUs

AMD ships Instella-MoE-16B-A3B, an open mixture-of-experts model trained on Instinct GPUs with 2.8B active parameters.

Inkling Small matches flagship skills at a quarter of the size

Thinking Machines shipped a quarter-size Inkling model that nears the flagship's benchmark scores.

Mira Murati’s Thinking Machines Lab debuts first open model

Inkling, a 975-billion-parameter open-weight model, lets organizations customize AI without relying on big labs.