llm-awq
by mit-han-lab
Pythonpushed almost 2 years ago
[MLSys 2024 Best Paper Award] AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration
AI summary
LLM Quantizer
An open-source software project that enables efficient and accurate low-bit weight quantization for large language models.
- stars
- 2.6K
- forks
- 212
- watching
- 25
- awesome list
- 1
Featured in 1 awesome list
Each link jumps to the spot where the list mentions llm-awq.