Implementation of the Llama architecture with RLHF + Q-learning - View it on GitHub
Star
158
Rank
177431