mistral-large-3

by Mistral

Mistral Large 3 is a MoE model with 67.5B total parameters and 41B active parameters, supporting a 256K-token context window. Trained from scratch on 3,000 NVIDIA H200 GPUs, it is one of the strongest permissively licensed open-weight models available. Designed for advanced reasoning and long-context understanding, Mistral Large 3 delivers performance on par with the best instruction-tuned open-weight models for general-purpose tasks, while also offering image understanding capabilities. Its multilingual strengths are particularly notable for non-English/Chinese languages, making it well-suited for global applications. Typical use cases include enterprise assistants, multilingual customer support, content generation and editing, data analysis over long documents, code assistance, and research workflows that require handling large corpora or complex instructions. With its MoE architecture, Mistral Large 3 balances strong performance with efficient inference, providing a versatile backbone for building reliable, production-grade AI systems.

API Pricing

Input$0.5 / 1M tokens
Output$1.5 / 1M tokens

Specifications

Context window256,000 tokens
Modalitiestext, image
Featuresfunction calling, structured outputs

FAQ

What is mistral-large-3?

Mistral Large 3 is a MoE model with 67.5B total parameters and 41B active parameters, supporting a 256K-token context window. Trained from scratch on 3,000 NVIDIA H200 GPUs, it is one of the strongest permissively licensed open-weight models available. Designed for advanced reasoning and long-context understanding, Mistral Large 3 delivers performance on par with the best instruction-tuned open-weight models for general-purpose tasks, while also offering image understanding capabilities. Its multilingual strengths are particularly notable for non-English/Chinese languages, making it well-suited for global applications. Typical use cases include enterprise assistants, multilingual customer support, content generation and editing, data analysis over long documents, code assistance, and research workflows that require handling large corpora or complex instructions. With its MoE architecture, Mistral Large 3 balances strong performance with efficient inference, providing a versatile backbone for building reliable, production-grade AI systems.

What is the context length of mistral-large-3?

mistral-large-3 has a 256,000 token context window.

How much does mistral-large-3 cost?

On AIHubMix, mistral-large-3 costs $0.5 per million input tokens and $1.5 per million output tokens.

What modalities does mistral-large-3 support?

mistral-large-3 accepts text and image input.

What features does mistral-large-3 support?

mistral-large-3 supports function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.

How do I call mistral-large-3 via API?

mistral-large-3 is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to mistral-large-3 — no other code changes needed.

Who develops mistral-large-3?

mistral-large-3 is developed by Mistral. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More from Mistral

AiHubmix-mistral-medium

by Mistral

Mistral Medium 3 is a SOTA & versatile model designed for a wide range of tasks…

$0.4/1M in · $2/1M out

chutesai/Mistral-Small-3.1-24B-Instruct-2503

by Mistral

Mistral's latest open-source small model; provided by chutes.ai.

$0.2/1M in · $0.8/1M out

codestral-latest

by Mistral

Mistral has launched a new code model - Codestral 25.01…

$0.4/1M in · $1.2/1M out

aihubmix-Mistral-Large-2411

by Mistral

The latest Mistral Large 2 model is deployed on Azure.

$2/1M in · $6/1M out

aihubmix-Mistral-large-2407

by Mistral
$3/1M in · $9/1M out

Mistral-large-2407

by Mistral
$3/1M in · $9/1M out

Use mistral-large-3 via the AIHubMix unified API — one interface for every major LLM.