End-to-end LLM engineering pipeline: LoRA fine-tuning on Qwen3-8B/14B, FP8 quantization, vLLM concurrency benchmarking, and BFCL evaluation — built for finance function-calling on H100 - View it on GitHub
Star
1
Rank
6255410