BenchLeader

GPT-3.5 Turbo (older v0613)

GPT-3.5 Turbo (older v0613) is an OpenAI proprietary model. It is not yet ranked: no independent benchmark results so far, only pricing and metadata. At $1.25 per million tokens blended it is pricier than most ranked models. Output speed of 10 tokens per second puts it in the slowest quarter, with a first token in 0.6 s.

Blended price
$1.25/M
$1.00 in · $2.00 out
Output speed
10 tok/s
OpenRouter traffic, 7-day median; not yet measured by Artificial Analysis
First answer
0.56 s
Context
4k
Released
date not published by our sources

No independent quality benchmark results yet — only pricing and metadata.

Where to run it

Every provider serving this model through OpenRouter, with throughput and first-token latency measured on live traffic over the last 30 minutes and each provider’s own price. Purple marks the best in each column.

ProviderSpeedFirst tokenInput $/MOutput $/MContextQuantisation
Azure5 tok/s0.49 s$1.00$2.004k

What a task costs

Estimates from list price, output speed and time to first answer for the best configuration. “With caching” assumes three-quarters of the input is served from the prompt cache. Reasoning tokens are not modelled.

WorkloadTokens in / outCostWith cachingTime
Chat reply400 / 300$0.001032.1 s
Summarise a 30-page report12,000 / 600$0.0131.1 min
Code edit6,000 / 1,500$0.00902.6 min
Agentic coding session60,000 / 4,000$0.0687.0 min
Structured extraction2,000 / 200$0.002421.6 s

See also

Data as of 13 Sept 2026. Compare with another model.

Cite as: BenchLeader, “GPT-3.5 Turbo (older v0613): benchmarks, pricing, speed and rank”, https://www.benchleader.com/models/gpt-3-5-turbo-0613, data as of 13 Sept 2026.