The Artificial PostAccount
Researchers

Tri Dao

Efficient AIPrinceton UniversityOfficial profile

Reducing memory traffic and computation in sequence models.

Dao studies the relationship between learning algorithms and the hardware that runs them. FlashAttention and Mamba anchor this selection, connecting memory-efficient attention with sequence models designed to use computation and memory differently.

Selected work

10 papers

An independent editorial profile. Inclusion does not imply Council membership or endorsement. Research is collaborative; coauthorship does not imply sole credit.

Selection & sources