BenchLeader
AmazonOpen weights

Llama 3.1 70B Instruct (US)

Llama 3.1 70B Instruct (US) is an Amazon open-weights model, released 23 Jul 2024. It is not yet ranked: no independent benchmark results so far, only pricing and metadata. At $0.720 per million tokens blended it is mid-priced.

Blended price
$0.720/M
$0.720 in · $0.720 out
Output speed
First answer
Context
128k
Released
23 Jul 2024

Versions

Amazon has shipped 2 models under this name. Each is ranked on its own results; a newer version often has fewer results so far, which holds its index nearer the average until more arrive.

ModelReleasedIndexRank
Llama 3.3 70B Instruct (US)6 Dec 2024
Llama 3.1 70B Instruct (US)this page23 Jul 2024

No independent quality benchmark results yet — only pricing and metadata.

What a task costs

Estimates from list price, output speed and time to first answer for the best configuration. “With caching” assumes three-quarters of the input is served from the prompt cache. Reasoning tokens are not modelled.

WorkloadTokens in / outCostWith cachingTime
Chat reply400 / 300$0.0005
Summarise a 30-page report12,000 / 600$0.0091
Code edit6,000 / 1,500$0.0054
Agentic coding session60,000 / 4,000$0.046
Structured extraction2,000 / 200$0.0016

See also

Data as of 13 Sept 2026. Compare with another model.

Cite as: BenchLeader, “Llama 3.1 70B Instruct (US): benchmarks, pricing, speed and rank”, https://www.benchleader.com/models/us-meta-llama3-1-70b, data as of 13 Sept 2026.