DeepSeek-V4 Pro vs Mistral Large 3
Head-to-head technical comparison of intelligence, capabilities, and API pricing.
Which should you use?
DeepSeek-V4 Pro leads on the composite index by 2.6 points (91.8 vs 89.2).
Pick DeepSeek-V4 Pro when
- Higher composite index — 91.8 against 89.2, a gap of 2.6 points.
- Ahead on Arena Elo by 19 points (1449 vs 1430), which is outside the board's published confidence intervals.
- Stronger on MMLU-Pro: 91.8% against 89.2%.
- Cheaper input tokens — $0.435 against $0.5 per 1M, so Mistral Large 3 costs 1.1x as much to feed.
- Better value per point of measured capability: $0.005 per index point against $0.006.
Pick Mistral Large 3 when
On these measures Mistral Large 3 does not lead DeepSeek-V4 Pro anywhere, so pick it only for reasons outside this table — an existing integration, a region, or a contract.
Value per point of measured capability, at list input pricing: DeepSeek-V4 Pro $0.005 · Mistral Large 3 $0.006 per index point. Run your own token mix through the token cost calculator — a comparison at list price ignores caching and batch discounts, which move real bills more than this gap does.
Where these numbers come from
Figures as of 2026-09-14. Arena Elo and pricing are read from the public board and each vendor's own pricing page; see the full leaderboard for all 34 models and which columns are measured rather than estimated, or the benchmark matrix for scores benchmark by benchmark.