Mistral Large 4 packs 1.05T parameters into an open-weight model

Mistral AI has released Mistral Large 4, an open-weight multimodal model built on a granular Mixture-of-Experts architecture. It has 49B active parameters, 1.05T total parameters, and a 1.6B vision encoder, with a 1M-token context window. Pricing starts at $0.68 per million input tokens and $2.09 per million output tokens, with cached input at $0.07. It supports structured outputs, function calling, document QnA, and agents.

Mistral Large 4 is a state-of-the-art, open-weight, general-purpose multimodal model with a granular Mixture-of-Experts architecture.

More from this day

2026-10-06