The Artificial PostAccount
All papers
Quantization · ORIGINAL RESEARCH

AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration

AWQ uses activation information to choose weight quantization strategies that preserve model performance under compression.

Opening the paper…

KEEP FOLLOWING THE IDEA

Meet the researchers.

Song Han