Mistral previews Large 4, a 1-trillion-parameter open-weight model built in Europe
Mistral opened a hosted preview of Large 4 on October 6: a multimodal mixture-of-experts model with 1.05 trillion parameters, a 1-million-token context window and open weights promised before the end of the month.
- 1.05T total parameters, about 49B active per token (mixture-of-experts)
- 1M-token context window; 1.6B-parameter vision encoder
- Trained on about 3,800 Nvidia Grace Blackwell GPUs in EU data centres; 160+ languages