brain-research/mirage-rl-stein

brain-research

Fetched on 2026/03/01 23:03

Fork of https://github.com/DartML/PPO-Stein-Control-Variate with modifications for the paper "The Mirage of Action-Dependent Baselines in Reinforcement Learning". - View it on GitHub

Star

Rank

14121396

brain-research

brain-research / mirage-rl-stein