Low-bit LLM inference on CPU/NPU with lookup table - View it on GitHub
Star
962
Rank
43964