BenchLeader
Zhipu AIReasoning modelAuto-detected

GLM-5-Turbo

Best configuration ranks #130 of 610 on the BenchLeader Index at 58.1 ±4.1. Released 16 Mar 2026.

Blended price
$1.90/M
$1.20 in · $4.00 out
Output speed
13 tok/s
First answer
6.58 s
Context
200k
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Overall index58
  2. Agents & tools63
  3. Knowledge56
  4. Instruction following71
  5. Long context62
  6. Composite63

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Reasoning

BenchmarkdefaultSource
GPQA Diamond (AA)not in index84.8%#132Artificial Analysis
Humanity's Last Exam (AA)not in index27.8%#125Artificial Analysis

Agents & tools

Knowledge

BenchmarkdefaultSource
AA-Omniscience-16.4#167Artificial Analysis

Instruction following

BenchmarkdefaultSource
IFBench73.2%#43Artificial Analysis

Long context

BenchmarkdefaultSource
AA-LCR71.7%#144Artificial Analysis

Composite

BenchmarkdefaultSource
AA Intelligence Index26.6#104Artificial Analysis

Where to run it

Every provider serving this model through OpenRouter, with throughput and first-token latency measured on live traffic over the last 30 minutes and each provider’s own price. Purple marks the best in each column.

ProviderSpeedFirst tokenInput $/MOutput $/MContextQuantisation
Z.ai13 tok/s6.58 s$1.20$4.00203k

Price history

Listed price per 1M tokens over time, as recorded by OpenRouter for the provider with the longest history.

$0.00$1.10$2.20$3.30$4.40Mar 26Apr 26May 26Jun 26Jul 26Aug 26
input outputnow $1.20 in · $4.00 out

What a task costs

Estimates from list price, output speed and time to first answer for the best configuration. “With caching” assumes three-quarters of the input is served from the prompt cache at $0.240 per 1M. Reasoning tokens are not modelled.

WorkloadTokens in / outCostWith cachingTime
Chat reply400 / 300$0.0017$0.001429.7 s
Summarise a 30-page report12,000 / 600$0.017$0.008252.7 s
Code edit6,000 / 1,500$0.013$0.00892.0 min
Agentic coding session60,000 / 4,000$0.088$0.0455.2 min
Structured extraction2,000 / 200$0.0032$0.001822.0 s

See also

Data as of 9 Sept 2026. Compare with another model.