A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce - View it on GitHub
Star
4
Rank
2858665