
Mistral Medium 3 — Benchmarks, Capabilities, and Use Cases
Mistral Medium 3 is a proprietary AI model from Mistral, classified as Advanced tier with a 128K tokens context window. Explore benchmark scores, strengths, and...
Mistral Large 3 is an advanced AI model suited for demanding production workloads, developed by Mistral and released in December 2025. As an open-weight model, its trained weights are publicly available — you can self-host it, fine-tune it on proprietary data, or run it on-premise without API dependencies. It supports a 256K-token context window, adequate for most documents and multi-turn applications.
On the FlowHunt AI Leaderboard it is one of 24 tracked models, evaluated across 3 benchmarks. Standout scores: 93.6% on MATH-500 (#1 of 4); 73.1% on MMLU-Pro (#5 of 8); 43.9% on GPQA Diamond (#14 of 14).
Mistral Large 3 was released on December 2, 2025 as Mistral’s most capable commercial model and, critically, as an open-weight release under the Apache 2.0 license. At 41B active parameters from a 675B total MoE architecture, it is the largest and most capable fully Apache-licensed frontier-adjacent model available. Mistral designed Large 3 to address European enterprise needs specifically: deployable in EU data centres, compliant with GDPR by architecture (data never leaves the deployment region), and usable for commercial applications without the restrictive licensing terms of many open-weight models. Its MMLU-Pro score of 84.2% and solid science reasoning profile have made it the default choice for European regulated industries — finance, healthcare, and legal — that need strong AI capability without sending data to US-based API providers.
Note: Enterprise generalist, Apache 2.0. Weak hard reasoning (GPQA 43.9%).
Rank among all 24 models tracked on this leaderboard. SR = self-reported · 3P / AA / C = independently verified.
| Benchmark | Score | Rank | Source | What it measures |
|---|---|---|---|---|
| GPQA Diamond | 43.9% | 14 of 14 | 3P | Graduate-level Google-proof science Q&A |
| MMLU-Pro | 73.1% | 5 of 8 | 3P | Hard knowledge reasoning, 12K questions |
| MATH-500 | 93.6% | 1 of 4 | 3P | Hendrycks math (500 problems) |
Head-to-head benchmark comparison with other Advanced-tier models. Higher is better for all metrics.
| Benchmark | Mistral Large 3 | Claude Sonnet 4.6 | GPT-4.1 |
|---|---|---|---|
| GPQA ◇ | 43.9% | — | — |
| MMLU-P | 73.1% | — | — |
| MATH | 93.6% | 89.0% | — |
Mistral AI is a French AI company founded in April 2023 by Arthur Mensch (formerly of Google DeepMind), Guillaume Lample, and Timothée Lacroix (both formerly of Meta AI Research). The company positions itself as a European champion in generative AI, combining a commercial API business with strategic open-weight releases. Mistral's flagship Large series and Apache 2.0-licensed models have made it the most prominent non-US, non-Chinese provider on most AI leaderboards. The company maintains EU headquarters in Paris, emphasises GDPR and EU regulatory compliance, and has articulated a distinct regulatory philosophy calling for lighter regulation of open-weight models compared to closed API systems.
Use Cases
Mistral Large 3 achieves 43% on GPQA Diamond (#14 of 14 models). It handles general scientific Q&A at a solid level, though frontier models score 25+ percentage points higher on the hardest graduate-level problems.
Mistral Large 3 is open-weight (41B/675B MoE parameters) — its trained weights are publicly available for download and self-deployment. Run it on your own GPU hardware or private cloud to ensure data never leaves your environment, eliminate per-token API costs at scale, or fine-tune the model on proprietary datasets.
More to Compare
Compare all 24 models across 11 benchmarks — sortable, sourced, and updated June 2026.

Mistral Medium 3 is a proprietary AI model from Mistral, classified as Advanced tier with a 128K tokens context window. Explore benchmark scores, strengths, and...

MiniMax M3 is a open-weight AI model from MiniMax (MoE (undisclosed)), classified as Frontier tier with a 1M tokens context window. It is available directly in ...

Mixtral 8x7B is a open-weight AI model from Mistral (12.9B/46.7B MoE), classified as Legacy tier with a 32K tokens context window. It is available directly in F...
Cookie Consent
We use cookies to enhance your browsing experience and analyze our traffic. See our privacy policy.