Preprint2021
A minimalist approach to offline reinforcement learning
Unknown
This paper proposes a minimalist offline RL approach that avoids out-of-distribution value overestimation by constraining the policy to the data support, achieving strong performance with simple modifications.
0Jan 1, 2021Reinforcement Learning