The Artificial PostAccount
All papers
Reasoning · ORIGINAL RESEARCH

HealthBench: Evaluating Large Language Models Towards Improved Human Health

Selected research in reasoning. Read the full paper, including the methods, experiments, and reported results.

Opening the paper…

KEEP FOLLOWING THE IDEA

Meet the researchers.

Jason Wei