Introducing Mistral Large 4
Officially called Mistral Large 4, it is the company’s largest and most capable model to date, combining instruction, reasoning, and agentic workflows in one system. The model outperforms leading open-weight competitors in coding, cybersecurity, finance, and legal tasks, while also excelling in multimodal reasoning—including visual grounding, satellite imagery analysis, and technical drawing…
Key points
- Trained on **3,800 NVIDIA Grace Blackwell GPUs** in Europe, with multilingual data across **160+ languages**, including all EU official languages
- Preview API available now via **Mistral Studio**; weights drop by **end of October**, with red-team testing by cybersecurity experts
Le Chonk is trained on 3,800 NVIDIA Grace Blackwell GPUs in Mistral’s European datacenters, using multilingual data across 160+ languages, including all EU official languages. It is currently available via a preview API on Mistral Studio, with weights scheduled for release by the end of October. The model is being red-teamed by cybersecurity leaders, vetted partners, and state authorities before full deployment. Mistral emphasizes its open-weight and self-deployment capabilities, positioning it as a sovereign AI solution for industries like cybersecurity, where provider-level refusals could block critical vulnerability research. The company also highlights its Reinforcement Learning (RL) at scale methodology, which adapts to the model’s evolving capabilities and supports complex, long-horizon tasks across domains.
Mistral claims le Chonk is state-of-the-art in cybersecurity (ranking top five globally on the Artificial Analysis Cyber Index), coding (leading open-weight models on DeepSWE v1.1 and SWE-Atlas-QnA), and multimodal reasoning (surpassing GPT-6 Astra on Dense 200). It also excels in financial analysis (outperforming GPT-6 Astra on HarveyAI’s Legal Agent benchmark) and scientific workflows (state-of-the-art on SciCode-Verified). The model’s resistance to prompt injections and refusal of malicious cybersecurity requests are also highlighted, with a 93.3% success rate on the Lakera B3 AI Security Benchmark.
Model pages: Mistral Large 4 → · GPT-6 Astra → · Claude Opus 5.5 →
The story so far
5 episodes →- Introducing Mistral Large 4this story
Introducing Mistral Large 4
Mistral AI · 6 October 2026
Loading the full article…
This text was published by Mistral AI. It is reproduced here with attribution so you can read it in full; the rights remain with the publisher. Read it at the source ↗
Coverage and discussion
5sources- Hacker News discussion · 5 pointsnews.ycombinator.com
- Mistral Unveils 1.05T-Parameter ‘Le Chonk’ MoE Model in Public PreviewPress · Unite.AI ·
- Mistral Says Its New AI Model ‘Le Chonk’ Is the Best Open-Weight Offering Outside of ChinaPress · Wired AI ·
- Mistral releases Mistral Large 4, dubbed "le Chonk", a 1T-parameter open-weight model for general agentic capabilities, trained on 4,000 Grace Blackwell GPUsPress · thedeepview.com ·
- Mistral unveils new AI model it says rivals best open systems from ChinaPress · CNBC Technology ·
The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us. Published by Martin K., who runs Digest AI and handles corrections.
More in Generative AI & Models
All →- OpenAI boosts GPT-6 speed by 50% in 28 Days update · 1 src
- Reflection AI debuts Beam, a 501B open-weight model rivaling GLM 5.2 with 3-4x less compute · 10 src
- Author tests four AI assistants on flawed forecasting scenarios · 1 src
- Anthropic prompts users to share voice data to enhance AI models · 8 src
- Anthropic cuts cache-reading fees by 75% in Claude Fable 5.1 · 1 src
Comments
via GitHub Discussions