The Artificial PostAccount
Researchers

Amanda Askell

AI safety

Studying helpfulness, honesty, and harmlessness in AI assistants.

Askell’s coauthored research examines how AI assistants should behave and how that behavior can be evaluated. These papers investigate helpfulness, honesty, harmlessness, principles for alignment, and failures that can survive conventional safety training.

Selected work

10 papers

An independent editorial profile. Inclusion does not imply Council membership or endorsement. Research is collaborative; coauthorship does not imply sole credit.

Selection & sources