The Artificial PostAccount
All papers
Efficient AI · ORIGINAL RESEARCH

Deep Compression: Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding

Deep Compression combines pruning, weight quantization, and coding to reduce the storage requirements of neural networks.

Opening the paper…

KEEP FOLLOWING THE IDEA

Meet the researchers.

Song Han