A framework for evaluating Large Language Models (LLMs) through strategic game-playing. - View it on GitHub
Star
2
Rank
4264962