Mistral Large 4: Europe's First Trillion-Parameter Model
Mistral released the public preview of Large 4, a 1-trillion-parameter sparse mixture-of-experts model trained on 3,800 Grace Blackwell GPUs in European data centers. Full open weights are scheduled for end of October under a custom licence.
- 1T total params, 49B active per token; trained from scratch in Europe on Grace Blackwell GPUs
- Targets cybersecurity, legal, and finance with claimed state-of-the-art open-weight performance
- Full weights release planned for end of October; preview API available now