BenchLeader
Mistral AIOpen weightsAuto-detected

Pixtral Large (latest)

Best configuration ranks #468 of 610 on the BenchLeader Index at 42.2 ±1.4. Last measured 27 Aug 2026. Released 1 Nov 2024.

Blended price
$3.00/M
$2.00 in · $6.00 out
Output speed
First answer
Context
128k
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Overall index42
  2. Instruction following38
  3. Multimodal37
  4. Composite38

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Reasoning

BenchmarkdefaultSource
GPQA Diamond (AA)not in index50.5%#405Artificial Analysis
Humanity's Last Exam (AA)not in index2.8%#518Artificial Analysis

Agents & tools

BenchmarkdefaultSource
τ²-Bench Telecom (AA)not in index36.5%#223Artificial Analysis

Instruction following

BenchmarkdefaultSource
IFBench34.5%#302Artificial Analysis

Multimodal

BenchmarkdefaultSource
LMArena Vision1089#110LMArena
MMMU-Pro50.6%#204Artificial Analysis
VISTA33.9%#47Scale AI SEAL

Composite

BenchmarkdefaultSource
AA Intelligence Index7.2#410Artificial Analysis

What a task costs

Estimates from list price, output speed and time to first answer for the best configuration. “With caching” assumes three-quarters of the input is served from the prompt cache. Reasoning tokens are not modelled.

WorkloadTokens in / outCostWith cachingTime
Chat reply400 / 300$0.0026
Summarise a 30-page report12,000 / 600$0.028
Code edit6,000 / 1,500$0.021
Agentic coding session60,000 / 4,000$0.144
Structured extraction2,000 / 200$0.0052

See also

Data as of 9 Sept 2026. Compare with another model.