The Artificial PostAccount
All papers
Benchmarks · ORIGINAL RESEARCH

Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

BIG-bench collects a broad set of language-model evaluation tasks to study capabilities and limitations beyond a single benchmark.

Opening the paper…

KEEP FOLLOWING THE IDEA

Meet the researchers.

Jared KaplanAmanda AskellChristopher D. ManningPercy LiangStefano ErmonDanqi ChenJason WeiYejin Choi