The Artificial PostAccount
Researchers

Noam Shazeer

Language models

Combining attention and sparse conditional computation.

Shazeer’s work connects attention-based architectures with sparse conditional computation. This selection includes the transformer and mixture-of-experts research, following approaches that increase model capacity while controlling the computation used for each input.

Selected work

10 papers

2017 · Transformers

Attention Is All You Need

The transformer replaces recurrence with attention. It became a foundation for language models and many other systems that process sequences.

Paper & context
2025 · AI for science

Gemma 3 Technical Report

Selected research in ai for science, efficient ai. Read the full paper, including the methods, experiments, and reported results.

Paper & context

An independent editorial profile. Inclusion does not imply Council membership or endorsement. Research is collaborative; coauthorship does not imply sole credit.

Selection & sources