JAX implementation of Adaptive Approximate Policy Iteration (Hao et al., 2021)
-
Updated
May 30, 2021 - Python
JAX implementation of Adaptive Approximate Policy Iteration (Hao et al., 2021)
Markov Decision Process DQN with Noisy Networks for Exploration (ICLR 2018) - 21.1% performance improvement over ε-greedy.
To associate your repository with the efficient-exploration topic, visit your repo's landing page and select "manage topics."