Model

Mistral Large 3 NO.24

Europe's flagship open-weight MoE — a strong, cheap, multilingual generalist.

Mistral AI · Mistral · Open · Open weights

Specification

Fig.01Mistral — Mistral AI

FRANCE

Context window
262Ktokens
Max output
33Ktokens
Input price
$0.50/ 1M
Output price
$1.50/ 1M
Throughput
Class
Open
Modalities
Text, Image, Pdf
Weights
Open
Parameters
675B
License
Apache 2.0
Knowledge cutoff
Released
Dec 2025

Overview

Mistral Large 3 (2512) is Mistral's open-weight flagship, a 675B-parameter sparse MoE with 41B active parameters under a permissive Apache 2.0 license. It pairs a 256K context, text+image input, and best-in-class European-language coverage at a low price. It is genuinely excellent value for multilingual and enterprise work, but sits below the Claude/GPT frontier on long-horizon agentic coding reliability.

Strengths

  • Best-in-class multilingual (esp. European languages)
  • Fully open weights under Apache 2.0
  • Aggressive pricing for a flagship-class model
  • 256K context with vision input

Trade-offs

  • Agentic coding reliability below the frontier
  • Optimized for breadth over deep reasoning
  • Vision capability is secondary
  • Large to self-host despite sparse activation

Fit

Best for

  • Multilingual and European-market apps
  • Enterprise/on-prem deployments
  • General reasoning and writing
  • Data-residency-constrained workloads

Not ideal for

  • Long autonomous coding agents
  • Hardest math/science reasoning (use Magistral)

Capabilities

Ten axes, normalized 0–100 and scored the same way across the whole catalog — so a 78 here means what a 78 means anywhere else on the bench.

Fig.02Capability radar

0 — 100

Reasoning
74
Coding
70
Math
72
Writing
78
Knowledge
80
Speed
66
Agentic
66
Vision
60
Multilingual
86
Long Context
76

Benchmarks

Public results, with independent third-party runs marked. Bars normalize percentages against 100 and Elo ratings against a 1500 ceiling.

Independent index

Artificial Analysis Intelligence Index

Composite of ~9–10 independent evals · Artificial Analysis

16/ 100

BenchmarkResult
Artificial Analysis Intelligence IndexIndependentGeneral · Artificial Analysis
16%
τ-benchIndependentAgentic · τ-bench
70.2%
LMArena EloIndependentGeneral · LMArena
1418Elo

Alternatives

Models in roughly the same class — the ones worth weighing against this record.

Index

All models/EDUCATION