by Mistral
Mistral Large 3 is a MoE model with 67.5B total parameters and 41B active parameters, supporting a 256K-token context window. Trained from scratch on 3,000 NVIDIA H200 GPUs, it is one of the strongest permissively licensed open-weight models available. Designed for advanced reasoning and long-context understanding, Mistral Large 3 delivers performance on par with the best instruction-tuned open-weight models for general-purpose tasks, while also offering image understanding capabilities. Its multilingual strengths are particularly notable for non-English/Chinese languages, making it well-suited for global applications. Typical use cases include enterprise assistants, multilingual customer support, content generation and editing, data analysis over long documents, code assistance, and research workflows that require handling large corpora or complex instructions. With its MoE architecture, Mistral Large 3 balances strong performance with efficient inference, providing a versatile backbone for building reliable, production-grade AI systems.
Mistral Large 3 is a MoE model with 67.5B total parameters and 41B active parameters, supporting a 256K-token context window. Trained from scratch on 3,000 NVIDIA H200 GPUs, it is one of the strongest permissively licensed open-weight models available. Designed for advanced reasoning and long-context understanding, Mistral Large 3 delivers performance on par with the best instruction-tuned open-weight models for general-purpose tasks, while also offering image understanding capabilities. Its multilingual strengths are particularly notable for non-English/Chinese languages, making it well-suited for global applications. Typical use cases include enterprise assistants, multilingual customer support, content generation and editing, data analysis over long documents, code assistance, and research workflows that require handling large corpora or complex instructions. With its MoE architecture, Mistral Large 3 balances strong performance with efficient inference, providing a versatile backbone for building reliable, production-grade AI systems.
mistral-large-3 has a 256,000 token context window.
On AIHubMix, mistral-large-3 costs $0.5 per million input tokens and $1.5 per million output tokens.
mistral-large-3 accepts text and image input.
mistral-large-3 supports function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.
mistral-large-3 is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to mistral-large-3 — no other code changes needed.
mistral-large-3 is developed by Mistral. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Mistral Medium 3 is a SOTA & versatile model designed for a wide range of tasks…
Mistral's latest open-source small model; provided by chutes.ai.
Mistral has launched a new code model - Codestral 25.01…
The latest Mistral Large 2 model is deployed on Azure.
Use mistral-large-3 via the AIHubMix unified API — one interface for every major LLM.