BenchLeader
Mistral AIAuto-detected

Mistral Large

Mistral Large is a Mistral AI proprietary model, released 26 Feb 2024. Its best configuration ranks #514 of 372 on the BenchLeader Index at 40.8 ±4.6, in the lower half. It scores highest in coding (41) and lowest in maths (30). At $6.00 per million tokens blended it is among the most expensive ranked models. Last measured 2 Sept 2026.

Blended price
$6.00/M
$4.00 in · $12.00 out
Output speed
First answer
Context
32k
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Overall index41
  2. Reasoning34
  3. Coding41
  4. Maths30
  5. Human preference39

Versions

Mistral AI has shipped 3 models under this name. Each is ranked on its own results; a newer version often has fewer results so far, which holds its index nearer the average until more arrive.

ModelReleasedIndexRank
Mistral Largethis page26 Feb 202440.8#514
Mistral Large 346.8#368
Mistral Large 239.5#540

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Reasoning

Coding

BenchmarkdefaultSource
LMArena Coding1295#265LMArena

Human preference

BenchmarkdefaultSource
LMArena Text1242#276LMArena

What a task costs

Estimates from list price, output speed and time to first answer for the best configuration. “With caching” assumes three-quarters of the input is served from the prompt cache. Reasoning tokens are not modelled.

WorkloadTokens in / outCostWith cachingTime
Chat reply400 / 300$0.0052
Summarise a 30-page report12,000 / 600$0.055
Code edit6,000 / 1,500$0.042
Agentic coding session60,000 / 4,000$0.288
Structured extraction2,000 / 200$0.010

See also

Data as of 10 Sept 2026. Compare with another model.