The Artificial PostAccount
All papers
AI safety · ORIGINAL RESEARCH

Discovering Language Model Behaviors with Model-Written Evaluations

Selected research in ai safety, scaling laws. Read the full paper, including the methods, experiments, and reported results.

Opening the paper…

KEEP FOLLOWING THE IDEA

Meet the researchers.

Dario AmodeiChris Olah