Nemotron 3 Ultra

NVIDIA's Nemotron 3 Ultra is a 550B hybrid Mamba-MoE built for long-running agents, the most intelligent US open-weights model, and NVIDIA shipped the data and recipes too, under a permissive license. What it is, what is open, and why you still cannot run it at home.

Mistral Large 3

Mistral Large 3 is a 675B Mixture-of-Experts flagship under Apache 2.0, Mistral taking its best model out from behind a research-only license. The licensing is the story; the benchmarks are good, not frontier. Here is the honest version.

GLM-5.2

Z.ai's GLM-5.2 is the first MIT-licensed model you can self-host that beats GPT-5.5 on long-horizon agentic coding, with a real 1M-token context, at about one-sixth the API cost. What changed, the numbers with their asterisks, and how to run it.