BenchLeader

Qwen3.8-Flash-Next

Best configuration ranks #108 of 610 on the BenchLeader Index at 59.3 ±7.4. Last measured 8 Sept 2026. Released 26 Aug 2026.

Blended price
$0.230/M
$0.150 in · $0.470 out
Output speed
64 tok/s
First answer
34 s
Context
256k
How it scores by categoryDashed line = average model (50). One step of 15 = one standard deviation.
  1. Overall index59
  2. Coding71
  3. Agents & tools49
  4. Knowledge59
  5. Multimodal64
  6. Long context66
  7. Composite69

Benchmark results

One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.

Reasoning

BenchmarkdefaultSourceTrend
LiveBench Reasoningnot in index87.4%#20LiveBench
GPQA Diamond (AA)not in index92.3%#36Artificial Analysis
Humanity's Last Exam (AA)not in index38.0%#66Artificial Analysis

Coding

BenchmarkdefaultSourceTrend
LMArena WebDev1631#8LMArena
LiveBench Codingnot in index72.5%#42LiveBench
SciCode (AA)not in index50.6%#63Artificial Analysis

Agents & tools

BenchmarkdefaultSourceTrend
LMArena Agent0.8#24LMArena
LiveBench Agentic Codingnot in index61.6%#8LiveBench

Maths

BenchmarkdefaultSourceTrend
LiveBench Mathematicsnot in index85.8%#39LiveBench

Knowledge

BenchmarkdefaultSourceTrend
LiveBench Data Analysisnot in index74.2%#29LiveBench
AA-Omniscience-9.7#136Artificial Analysis

Instruction following

BenchmarkdefaultSourceTrend
LiveBench Languagenot in index74.6%#40LiveBench

Multimodal

BenchmarkdefaultSourceTrend
MMMU-Pro79.8%#40Artificial Analysis

Long context

BenchmarkdefaultSourceTrend
AA-LCR79.7%#59Artificial Analysis

Composite

BenchmarkdefaultSourceTrend
LiveBench76.2%#22LiveBench
AA Intelligence Index42.2#28Artificial Analysis

What a task costs

Estimates from list price, output speed and time to first answer for the best configuration. “With caching” assumes three-quarters of the input is served from the prompt cache. Reasoning tokens are not modelled.

WorkloadTokens in / outCostWith cachingTime
Chat reply400 / 300$0.000238.9 s
Summarise a 30-page report12,000 / 600$0.002143.6 s
Code edit6,000 / 1,500$0.001657.8 s
Agentic coding session60,000 / 4,000$0.0111.6 min
Structured extraction2,000 / 200$0.000437.3 s

See also

Data as of 9 Sept 2026. Compare with another model.