Mistral Large 3
A solid pick for self-hosting and full data control.
Overview
Mistral Large 3 is THE reference model for sovereign European use: France-hosted, full GDPR compliance, French team. French quality is excellent (Claude-level). Solid coding and reasoning performance, slightly behind the priciest US frontiers but at a much lower price.
Skill profile
What do these scores mean?
PhD-level science questions (physics, chemistry, biology), with no tool access.
Resolving real GitHub issues under real conditions (SWE-bench Verified).
Competition-level math problems.
General knowledge across dozens of academic subjects.
Python code generation from specifications.
Strengths
- Excellent at code
- Strong French quality
- Has a free tier
- Open-source and self-hostable
Who is it for
- you want to control cost or self-host
- you work in French
- you have GDPR constraints
Ideal use cases
- GDPR-compliant European apps
- European multilingual
- French apps
- EU enterprise
Access & availability
Key specifications
Estimate your monthly cost
Per-token API pricingFor the same usage
- GPT-5$108+6%
- Claude Sonnet 4.6$196+91%
- Grok 4$256+150%
Indicative estimate based on standard API rates (excluding caching, batch and volume discounts). Always check the official pricing before committing.
Privacy
Run it locally
| Quantization | Disk | RAM / VRAM | Typical hardware |
|---|---|---|---|
| Q4 · recommended | 384.7 GB | 425 GB | Server GPU infra (H100/A100…) |
| Q8 · balanced | 722.2 GB | 796 GB | Server GPU infra (H100/A100…) |
| FP16 · max quality | 1350 GB | 1487 GB | Server GPU infra (H100/A100…) |
Estimates for a moderate context. Long contexts need more RAM (KV cache).
Advanced data · for expertsArchitecture, modalities, detailed cost, full benchmarks▾
| Arena Elo | 1335 |
| MMLU | 84.5% |
| GPQA | 64.0% |
| HumanEval | 88.0% |
| SWE-Bench | 50.0% |
| MATH | 84.0% |