Pinned Loading jaxrl jaxrl Public template JAX (Flax) implementation of algorithms for Deep Reinforcement Learning with continuous action spaces. Jupyter Notebook 758 75 DrQ: Data regularized Q Jupyter Notebook 422 54 PyTorch implementation of Advantage Actor Critic (A2C), Proximal Policy Optimization (PPO), Scalable trust-region method for deep reinforcement learning using Kronecker-factored approximation (ACKT… Python 3.9k 840 PyTorch implementation of Asynchronous Advantage Actor Critic (A3C) from "Asynchronous Methods for Deep Reinforcement Learning". Python 1.3k 281 PyTorch implementations of algorithms for density estimation Python 589 75