[ICLR'26] MARSHAL: Incentivizing Multi-Agent Reasoning via Self-Play with Strategic LLMs - View it on GitHub
Star
56
Rank
491912