A PyTorch-native inference engine with cache, parallelism, quantization and cpu offload for DiTs. - View it on GitHub
Star
1248
Rank
35918