brain-research/mirage-rl-bpttv

brain-research

Fetched on 2026/03/01 23:03

Fork of https://github.com/wgrathwohl/BackpropThroughTheVoidRL with modifications for the paper "The Mirage of Action-Dependent Baselines in Reinforcement Learning". - View it on GitHub

Star

Rank

4265675

brain-research

brain-research / mirage-rl-bpttv