Fork of https://github.com/s…ixiang/rllabplusplus with modifications for the paper "The Mirage of Action-Dependent Baselines in Reinforcement Learning".
An updated Consolas for Powerline font. This works with the new vim-airline, too.
pytorch handbook是一本开源的书籍,目标是帮助那些希望和使用PyTorch进行深度学习开发和研究的朋友快速入门,其中包含的Pytorch教程全部通过测试保证可以成功运行
PyTorch0.4 implementation of: actor critic / proximal policy optimization / acer / ddpg / twin dueling ddpg / soft actor critic / generative adversarial imitation learning / hindsight experience replay
Pytorch solutions for UC Berkeley's cs285 assignments
Pytorch Implementation of DQN / DDQN / Prioritized replay/ noisy networks/ distributional values/ Rainbow/ hierarchical RL
OpenAI Baselines: high-quality implementations of reinforcement learning algorithms