PyTorch implementation of the Q-Learning Algorithm Normalized Advantage Function for continuous control problems + PER and N-step Method
reinforcement-learning q-learning dqn reinforcement-learning-algorithms continuous-control naf ddpg-algorithm prioritized-experience-replay normalized-advantage-functions q-learning-algorithm n-step-bootstrapping
-
Updated
Feb 16, 2021 - Jupyter Notebook