mistral.ai via Hacker News

Mistral Previews 1T-Param Open-Weight Large 4 'Le Chonk'

TL;DR

  • Mistral is previewing Mistral Large 4 (le Chonk), a 1-trillion-parameter mixture-of-experts with 49B active, trained on 3,800 Grace Blackwell GPUs.
  • Weights will ship by the end of October; preview API is priced at $1.36 per million input tokens and $4.18 per million output.
  • Mistral benchmarks ML4 against DeepSeek, Qwen, Kimi and GLM, and claims it beats any US or European open-weight model.

Mistral said on Monday it had trained a 1 trillion-parameter mixture-of-experts model, 49 billion active, on 3,800 Nvidia Grace Blackwell GPUs in its own European datacenters. The French lab is calling it Mistral Large 4, or in its own words, "Unofficially ML4, very officially: le Chonk."

A public preview went live on Mistral's API today, priced at $1.36 per million input tokens and $4.18 per million output. "We will release the weights by the end of the month," the company says. ML4 supports 160+ languages, including every official EU language.

The competitive frame is unusual. Mistral benchmarks ML4 head-to-head against DeepSeek V4 Pro, Qwen3.8 Max, Kimi K3 and GLM-5.3, and claims the model is "significantly outperforming any open-weight model developed in the US or Europe." The write-up pitches cybersecurity as a wedge, reporting 93% on Cybench and 82% on a vulnerability reproduction and patching test. The company notes that "several leading closed models, including Claude Opus 5.5 and GPT-6 Astra, score near zero on the same test because they refuse to perform the task."

"ML4 is the first milestone on the roadmap funded by our €3 billion Series D," Mistral writes, referring to the raise it calls the largest equity round ever by a European technology company. That round closed in September with Samsung leading. The post is signed "By Mistral"; no executive is quoted by name.