End-to-end pipeline for quantization-aware fine-tuning of Llama models with Docker, GCP integration, and Ollama API-based inference. - View it on GitHub
Star
1
Rank
6255410