The Artificial PostAccount
All papers
AI safety · ORIGINAL RESEARCH

The Capacity for Moral Self-Correction in Large Language Models

Selected research in ai safety, scaling laws. Read the full paper, including the methods, experiments, and reported results.

Opening the paper…

KEEP FOLLOWING THE IDEA

Meet the researchers.

Dario AmodeiChris Olah