Gitstar Ranking
Users
Organizations
Repositories
Rankings
Users
Organizations
Repositories
Sign in with GitHub
gmh5225
Fetched on 2026/08/06 12:23
gmh5225
/
exllama
A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights. -
View it on GitHub
Star
0
Rank
14132882