BenchLeader
DeepSeekOpen weightsReasoning model

DeepSeek R1 0528

DeepSeek R1 0528 is a DeepSeek open-weights reasoning model, released 28 May 2025. Its best configuration ranks #313 of 372 on the BenchLeader Index at 49.7 ±6.9, in the lower half. It scores highest in human preference (61) and lowest in knowledge (35). At $1.00 per million tokens blended it is mid-priced. Output speed of 24 tokens per second puts it in the slowest quarter, with a first token in 1.5 s. Last measured 2 Sept 2026.

Blended price
$1.00/M
$0.574 in · $2.29 out
Output speed
24 tok/s
OpenRouter traffic, 7-day median; not yet measured by Artificial Analysis
First answer
1.52 s
Context
164k
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Overall index50
  2. Reasoning52
  3. Coding53
  4. Maths48
  5. Knowledge35
  6. Human preference61

Versions

DeepSeek has shipped 2 models under this name. Each is ranked on its own results; a newer version often has fewer results so far, which holds its index nearer the average until more arrive.

ModelReleasedIndexRank
DeepSeek R1 0528this page28 May 202549.7#313
DeepSeek R120 Jan 202549.7#312

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Coding

BenchmarkdefaultSource
WeirdML41.6%#96WeirdML
LMArena Coding1464#108LMArena

Knowledge

Human preference

BenchmarkdefaultSource
LMArena Text1421#105LMArena

Where to run it

Every provider serving this model through OpenRouter, with throughput and first-token latency measured on live traffic over the last 30 minutes and each provider’s own price. Purple marks the best in each column.

ProviderSpeedFirst tokenInput $/MOutput $/MContextQuantisation
StreamLake52 tok/s3.28 s$0.571$2.29128k
SiliconFlow24 tok/s1.52 s$0.500$2.18164kfp8
DeepInfra17 tok/s0.86 s$0.500$2.15164kfp4

What a task costs

Estimates from list price, output speed and time to first answer for the best configuration. “With caching” assumes three-quarters of the input is served from the prompt cache at $0.350 per 1M. Reasoning tokens are not modelled.

WorkloadTokens in / outCostWith cachingTime
Chat reply400 / 300$0.0009$0.000914.0 s
Summarise a 30-page report12,000 / 600$0.0083$0.006226.5 s
Code edit6,000 / 1,500$0.0069$0.00591.1 min
Agentic coding session60,000 / 4,000$0.044$0.0342.8 min
Structured extraction2,000 / 200$0.0016$0.00139.9 s

See also

Data as of 10 Sept 2026. Compare with another model.