The Artificial PostAccount
Researchers

Song Han

Efficient AIMITOfficial profile

Compressing and accelerating neural networks for deployment.

Han studies how to make neural networks efficient enough to deploy. These papers connect pruning and quantization with memory-efficient language-model inference and efficient generative models, considering both algorithms and the hardware executing them.

Selected work

10 papers

An independent editorial profile. Inclusion does not imply Council membership or endorsement. Research is collaborative; coauthorship does not imply sole credit.

Selection & sources