Mistral logo

Mistral

Mistral Large 3

AdvancedOpen-weightAvailable in FlowHunt

Mistral Large 3 is an advanced AI model suited for demanding production workloads, developed by Mistral and released in December 2025. As an open-weight model, its trained weights are publicly available — you can self-host it, fine-tune it on proprietary data, or run it on-premise without API dependencies. It supports a 256K-token context window, adequate for most documents and multi-turn applications.

On the FlowHunt AI Leaderboard it is one of 24 tracked models, evaluated across 3 benchmarks. Standout scores: 93.6% on MATH-500 (#1 of 4); 73.1% on MMLU-Pro (#5 of 8); 43.9% on GPQA Diamond (#14 of 14).

Mistral Large 3 was released on December 2, 2025 as Mistral’s most capable commercial model and, critically, as an open-weight release under the Apache 2.0 license. At 41B active parameters from a 675B total MoE architecture, it is the largest and most capable fully Apache-licensed frontier-adjacent model available. Mistral designed Large 3 to address European enterprise needs specifically: deployable in EU data centres, compliant with GDPR by architecture (data never leaves the deployment region), and usable for commercial applications without the restrictive licensing terms of many open-weight models. Its MMLU-Pro score of 84.2% and solid science reasoning profile have made it the default choice for European regulated industries — finance, healthcare, and legal — that need strong AI capability without sending data to US-based API providers.

Note: Enterprise generalist, Apache 2.0. Weak hard reasoning (GPQA 43.9%).

Released
December 2025
Parameters
41B/675B MoE
Context
256K tokens
Weights
Open
Tier
Advanced
Provider
Mistral

Mistral Large 3 Benchmark Rankings

Rank among all 24 models tracked on this leaderboard. SR = self-reported · 3P / AA / C = independently verified.

BenchmarkScoreRankSourceWhat it measures
GPQA Diamond43.9%14 of 143PGraduate-level Google-proof science Q&A
MMLU-Pro73.1%5 of 83PHard knowledge reasoning, 12K questions
MATH-50093.6%1 of 43PHendrycks math (500 problems)

Detailed Benchmark Scores

GPQA ◇ 3P
43.9%
Graduate-level Google-proof science Q&A
MMLU-P 3P
73.1%
Hard knowledge reasoning, 12K questions
MATH 3P
93.6%
Hendrycks math (500 problems)

Mistral Large 3 vs. Advanced Peers

Head-to-head benchmark comparison with other Advanced-tier models. Higher is better for all metrics.

BenchmarkMistral Large 3Claude Sonnet 4.6GPT-4.1
GPQA ◇43.9%
MMLU-P73.1%
MATH93.6%89.0%

About Mistral

Mistral Founded 2023 · Paris, France
Website →

Mistral AI is a French AI company founded in April 2023 by Arthur Mensch (formerly of Google DeepMind), Guillaume Lample, and Timothée Lacroix (both formerly of Meta AI Research). The company positions itself as a European champion in generative AI, combining a commercial API business with strategic open-weight releases. Mistral's flagship Large series and Apache 2.0-licensed models have made it the most prominent non-US, non-Chinese provider on most AI leaderboards. The company maintains EU headquarters in Paris, emphasises GDPR and EU regulatory compliance, and has articulated a distinct regulatory philosophy calling for lighter regulation of open-weight models compared to closed API systems.

Use Cases

What to Use Mistral Large 3 For

Scientific & Technical Reasoning
Self-Hosted & Private Deployments

Strengths & Limitations

Strengths

  • Advanced mathematics — 93% MATH-500
  • Self-hosted & data-private deployments (open-weight)
  • Long-context tasks (256K context window)

Limitations

  • Below-frontier hard science reasoning (43% GPQA — frontier is 90%+)

Browse the Full AI Model Leaderboard

Frequently asked questions

Learn more

Mistral Medium 3 — Benchmarks, Capabilities, and Use Cases
Mistral Medium 3 — Benchmarks, Capabilities, and Use Cases

Mistral Medium 3 — Benchmarks, Capabilities, and Use Cases

Mistral Medium 3 is a proprietary AI model from Mistral, classified as Advanced tier with a 128K tokens context window. Explore benchmark scores, strengths, and...

2 min read
MiniMax M3 — Benchmarks, Capabilities, and Use Cases
MiniMax M3 — Benchmarks, Capabilities, and Use Cases

MiniMax M3 — Benchmarks, Capabilities, and Use Cases

MiniMax M3 is a open-weight AI model from MiniMax (MoE (undisclosed)), classified as Frontier tier with a 1M tokens context window. It is available directly in ...

4 min read
Mixtral 8x7B — Benchmarks, Capabilities, and Use Cases
Mixtral 8x7B — Benchmarks, Capabilities, and Use Cases

Mixtral 8x7B — Benchmarks, Capabilities, and Use Cases

Mixtral 8x7B is a open-weight AI model from Mistral (12.9B/46.7B MoE), classified as Legacy tier with a 32K tokens context window. It is available directly in F...

3 min read