LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar - View it on GitHub
Star
1
Rank
6084978