[ICLR2024 spotlight] OmniQuant is a simple and powerful quantization technique for LLMs. - View it on GitHub
Star
679
Rank
49520