BenchLeader

LAMBADA

Predicting the last word of a passage. A 2016 benchmark, saturated for current models; kept for history.

As of 19 Sept 2026, Megatron-Turing NLG 530B leads LAMBADA on BenchLeader with 87.2%, ahead of InstructGPT 175B at 86.4%, across 24 model configurations with a published result.

Published by
Epoch AI Benchmarking Hub
Category
Knowledge
Index weight
Reference only
Models
24
Data as of
19 Sept 2026

CC BY 4.0 — Epoch AI, ‘AI Benchmarking Hub’, epoch.ai/benchmarks. Mirrored boards credit their original publishers.

What the test looks like

Predicting the last word of a passage.

How it is scored

Accuracy, as compiled by Epoch AI from published results.

What to keep in mind

Saturated: frontier models score near the ceiling, so it separates only older or smaller models, and results are largely self-reported by developers.

24 of 24
#
1Megatron-Turing NLG 530BNVIDIA87.2%2021-10-11
2InstructGPT 175BOpenAI86.4%2022-01-27
3Falcon-180BTII79.8%2023-09-06
4Llama 2-70BMetaopen ↗78.9%34.92023-07-18
5Inflection-1Inflection AI78.5%2023-06-22
6Palm 540BGoogle77.9%2022-04-04
7LLaMA-65BMetaopen77.7%2023-02-24
8Falcon-40BTII77.3%2023-03-15
9LLaMA-33BMeta77.2%2023-02-24
10Llama 2-13BMetaopen ↗76.5%36.52023-07-18
11LLaMA-13BMeta75.2%2023-02-24
12Falcon-7BTII74.9%2023-04-24
13Baichuan2-13BBaichuan74.0%2023-09-06
14Llama 2-7BMetaopen ↗73.3%34.22023-07-18
15Baichuan 2-7BBaichuan73.3%2023-09-20
16LLaMA-7BMeta73.3%2023-02-24
17internlm-20bShanghai AI Lab71.8%2023-09-18
18Stable Beluga 2Stability AI71.3%2023-07-20
19Qwen-14BAlibaba71.1%2023-09-28
20MPT-7BMosaicML70.0%2023-05-05
21Qwen-7BAlibaba67.9%2023-09-28
22internlm-7bShanghai AI Lab67.0%2023-07-05
23Qwen-1_8BAlibaba58.4%2023-11-30
24chatglm2-6bZhipu AI54.3%2023-06-24

Cite as: BenchLeader, “LAMBADA leaderboard”, https://www.benchleader.com/benchmarks/lambada, data as of 19 Sept 2026.

LAMBADA: questions

What does LAMBADA measure?
Predicting the last word of a passage. Scores are reported in percent of tasks solved; higher is better.
Which AI model leads LAMBADA?
Megatron-Turing NLG 530B leads LAMBADA with 87.2% as of 19 Sept 2026, ahead of InstructGPT 175B at 86.4%.
How many models have LAMBADA results?
24 model configurations have a LAMBADA result on BenchLeader, all taken from Epoch AI Benchmarking Hub.
Who runs LAMBADA and how often is it updated?
LAMBADA is published by Epoch AI Benchmarking Hub. BenchLeader re-reads the published results every morning and records the date each result was published.
Does LAMBADA count toward the BenchLeader Index?
No. LAMBADA is shown for reference but left out of the composite index.