SpargeAttention: A training-free sparse attention that can accelerate any model inference. - View it on GitHub
Star
0
Rank
13221298