Mistral AI releases Mistral 3 family β Mistral Large 3 (675B MoE) under Apache 2.0
Mistral AI launched the Mistral 3 model family: flagship Mistral Large 3 (675B total / 41B active parameters, sparse MoE), plus Ministral 3B, 8B, and 14B models. All released under Apache 2.0. Large 3 was trained from scratch on ~3,000 H200 GPUs and debuted at #2 open-source non-reasoning model on LMArena. Key benchmark numbers: 85.5% MMLU (8-lang), 92% HumanEval, 93.6% MATH-500. Native vision support and 256K context window. Mistral also announced a multi-year cloud deal with HSBC.