Model

Magistral Medium 1.2 NO.27

Mistral's transparent-chain reasoning flagship, now multimodal.

Mistral AI · Mistral · Reasoning · Closed weights

Specification

Fig.01Mistral — Mistral AI

FRANCE

Context window
131Ktokens
Max output
41Ktokens
Input price
$2/ 1M
Output price
$5/ 1M
Throughput
Class
Reasoning
Modalities
Text, Image
Weights
Closed
Knowledge cutoff
Jun 2025
Released
Sep 2025

Overview

Magistral Medium 1.2 is Mistral's proprietary reasoning model, producing explicit multilingual chains of thought and strong performance on math and structured problem-solving, now with a vision encoder for reasoning over images. It is competitive on reasoning benchmarks and notably good at showing its work in the user's language, but its agentic coding reliability remains below the closed frontier. Best used for math, logic, and analysis rather than autonomous engineering.

Strengths

  • Dedicated, transparent reasoning traces
  • Strong competition-math performance (AIME)
  • Multimodal reasoning over text and images
  • Function calling and JSON mode support

Trade-offs

  • Slower and pricier per task (long thinking output)
  • Closed weights (Magistral Small is the open variant)
  • Agentic coding below the frontier
  • Narrower general knowledge than flagships

Fit

Best for

  • Complex math and scientific reasoning
  • Multistep logic and analysis
  • Visual reasoning tasks
  • Auditable chain-of-thought workflows

Not ideal for

  • Latency-sensitive or high-volume chat
  • Long autonomous coding agents

Capabilities

Ten axes, normalized 0–100 and scored the same way across the whole catalog — so a 78 here means what a 78 means anywhere else on the bench.

Fig.02Capability radar

0 — 100

Reasoning
80
Coding
66
Math
84
Writing
72
Knowledge
76
Speed
58
Agentic
62
Vision
60
Multilingual
80
Long Context
72

Benchmarks

Public results, with independent third-party runs marked. Bars normalize percentages against 100 and Elo ratings against a 1500 ceiling.

Independent index

Artificial Analysis Intelligence Index

Composite of ~9–10 independent evals · Artificial Analysis

18/ 100

BenchmarkResult
Artificial Analysis Intelligence IndexIndependentGeneral · Artificial Analysis
18%
Aider PolyglotIndependentCoding · Aider
47.1%

Alternatives

Models in roughly the same class — the ones worth weighing against this record.

Index

All models/EDUCATION