BenchLeader
AmazonOpen weights

Llama 3.1 8B Instruct (US)

Llama 3.1 8B Instruct (US) is an Amazon open-weights model, released 23 Jul 2024. It is not yet ranked: no independent benchmark results so far, only pricing and metadata. At $0.220 per million tokens blended it is cheaper than most ranked models.

Blended price
$0.220/M
$0.220 in · $0.220 out
Output speed
First answer
Context
128k
Released
23 Jul 2024

No independent quality benchmark results yet — only pricing and metadata.

What a task costs

Estimates from list price, output speed and time to first answer for the best configuration. “With caching” assumes three-quarters of the input is served from the prompt cache. Reasoning tokens are not modelled.

WorkloadTokens in / outCostWith cachingTime
Chat reply400 / 300$0.0002
Summarise a 30-page report12,000 / 600$0.0028
Code edit6,000 / 1,500$0.0017
Agentic coding session60,000 / 4,000$0.014
Structured extraction2,000 / 200$0.0005

See also

Data as of 13 Sept 2026. Compare with another model.

Cite as: BenchLeader, “Llama 3.1 8B Instruct (US): benchmarks, pricing, speed and rank”, https://www.benchleader.com/models/us-meta-llama3-1-8b, data as of 13 Sept 2026.