The Artificial PostAccount
All papers
Scaling laws · ORIGINAL RESEARCH

Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming

Selected research in scaling laws, ai safety. Read the full paper, including the methods, experiments, and reported results.

Opening the paper…

KEEP FOLLOWING THE IDEA

Meet the researchers.

Jared KaplanAmanda Askell