Live
Benchmarks

LAMBADA

Models scored
53
Best score
0.87
Metric
Score

LAMBADA is an AI evaluation tracked by Epoch AI, with 53 scored model versions on record. The scores here are reported by third parties rather than produced by Epoch AI running the evaluation itself.

Scored models span December 2021 to November 2023. The best score recorded is 0.87, measured as Score. Models come from MosaicML, Baichuan, Meta AI, Stability AI, Alibaba, OpenAI, DeepMind, Microsoft,NVIDIA and others.

Full record
Score metric
Score
Models scored
53
Best score recorded
0.87
Earliest model scored
Dec 8, 2021
Most recent model scored
Nov 30, 2023
Third-party reported
Yes
Leaderboard
01Megatron-Turing NLG 530BMicrosoft,NVIDIA0.87
02text-davinci-001OpenAI0.86
03falcon-180BTechnology Innovation Institute0.8
04Llama-2-70b-hfMeta AI0.79
05Inflection-1Inflection AI0.79
06PaLM 540BGoogle Research0.78
07LLaMA-65BMeta AI0.78
08Chinchilla (70B)DeepMind0.77
09falcon-40bTechnology Innovation Institute0.77
10LLaMA-33BMeta AI0.77
11Llama-2-13bMeta AI0.77
12LLaMA-13BMeta AI0.75
13falcon-7bTechnology Innovation Institute0.75
14Gopher (280B)DeepMind0.74
15Baichuan-2-13B-BaseBaichuan0.74
16Baichuan-2-7B-BaseBaichuan0.73
17LLaMA-7BMeta AI0.73
18Llama-2-7bMeta AI0.73
19internlm-20b0.72
20StableBeluga2Stability AI0.71
21Qwen-14BAlibaba0.71
22mpt-7bMosaicML0.7
23Qwen-7BAlibaba0.68
24internlm-7b0.67
25Qwen-1_8B0.58
26chatglm2-6b0.54
More benchmarks
SourceEpoch AI, 'AI Benchmarking Hub'. Published online at epoch.ai. Retrieved 2026-07-29 from https://epoch.ai/benchmarks. Licensed under CC BY 4.0.
← All benchmarks