Mistral AI Releases Mistral Large 4 (Le Chonk): A 1.05T Parameter Multimodal MoE Model
Quick Answer
Mistral AI has unveiled Mistral Large 4, also known as Le Chonk, a 1.05 trillion parameter multimodal Mixture of Experts model featuring 49 billion active parameters and a 1 million token context window.
Quick Take
The model, trained on 3,800 NVIDIA Grace Blackwell GPUs, is now available via API, with open weights expected by October 2026.
Key Points
- Mistral Large 4 features a total of 1.05 trillion parameters.
- The model has 49 billion active parameters for enhanced performance.
- It supports native image input and a context window of 1 million tokens.
- Trained on 3,800 NVIDIA Grace Blackwell GPUs in European datacenters.
- API access is live, with open weights scheduled for late October 2026.
Article Excerpt
From source RSS / original summaryMistral AI has released Mistral Large 4, nicknamed Le Chonk, as a public preview. It is a 1. 05 trillion parameter Mixture of Experts model with 49 billion active parameters, native image input, and a 1 million token context window, trained on 3,800 NVIDIA Grace Blackwell GPUs in Mistral's own European datacenters. The API is live now; open weights ship end of October 2026. The post Mistral AI Releases Mistral Large 4 (Le Chonk): A 1. 05T Parameter Multimodal MoE Model appeared first on MarkTechPost.
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from MarkTechPost
See more →Meet Flash-KMeans: An IO-Aware, Exact K-Means That Runs Over 200× Faster Than FAISS on GPUs
Flash-KMeans is an open-source, IO-aware k-means implementation that operates over 200× faster than FAISS on NVIDIA H200 GPUs. It achieves 17.9× end-to-end and 33× speedup over cuML by optimizing distance calculations and updating mechanisms without approximating results. This advancement significantly enhances performance for data scientists and machine learning practitioners.