awesome-rl
Reinforcement learning resources curated
https://github.com/aikorea/awesome-rl
Last synced: 18 days ago
JSON representation
-
Applications
-
Control
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
- [Paper
-
Game Playing
- [Paper
- [arXiv
- [arXiv
- [DOI - tw3IWgTseRnLpAc9xQq-vTA2Z5Ji9lg16_WvCy4SaOgpK5XXA6ecqo8d8J7l4EJsdjwai53GqKt-7JuioG0r3iV67MQIro74l6IxvmcVNKBgOwiMGi8U0izJStLpmQp6Vmi_8Lw_A%3D%3D) [[Code]](https://sites.google.com/a/deepmind.com/dqn/) [[Video]](https://www.youtube.com/watch?v=iqXKQf2BOSE)
- [Paper
- [DOI - 019-1724-z.epdf) [[Video]](https://deepmind.com/research/open-source/alphastar-resources)
- [arXiv
- [arXiv
- Flappy Bird Reinforcement Learning
- [Paper
-
Human Computer Interaction
-
Operations Research
-
Robotics
-
-
Codes
-
Human Computer Interaction
- Pole-Cart Problem
- Q-learning Controller
- MATLAB Environment and GUI for Reinforcement Learning
- Reinforcement Learning Repository - University of Massachusetts, Amherst
- Brown-UMBC Reinforcement Learning and Planning Library (Java)
- Reinforcement Learning in R (MDP, Value Iteration)
- Reinforcement Learning Environment in Python and MATLAB
- PyBrain Library - Python-Based Reinforcement learning, Artificial intelligence, and Neural network
- RLPy Framework - Value-Function-Based Reinforcement Learning Framework for Education and Research
- Maja - Machine learning framework for problems in Reinforcement Learning in python
- TeachingBox - Java based Reinforcement Learning framework
- Policy Gradient Reinforcement Learning Toolbox for MATLAB
- PIQLE - Platform Implementing Q-Learning and other RL algorithms
- BeliefBox - Bayesian reinforcement learning library and toolkit
- MATLAB Code
- POMDP for Dummies
- MATLAB Software, presentations, and demo videos
-
- MATLAB Code (BROKEN LINK)
- C/Lisp Code
- Book
- Python Code
- Julia Code
- Exercise Solutions
- C/Lisp Code
- Brown-UMBC Reinforcement Learning and Planning Library (Java)
- PyBrain Library - Python-Based Reinforcement learning, Artificial intelligence, and Neural network
- Maja - Machine learning framework for problems in Reinforcement Learning in python
- TeachingBox - Java based Reinforcement Learning framework
- Policy Gradient Reinforcement Learning Toolbox for MATLAB
- Deep Q-Learning with TensorFlow - A deep Q learning demonstration using Google Tensorflow
- Atari - Deep Q-networks and asynchronous agents in Torch
- AgentNet - A python library for deep reinforcement learning and custom recurrent networks using Theano+Lasagne.
- Reinforcement Learning Examples by RLCode - A Collection of minimal and clean reinforcement learning examples
- OpenAI Baselines - Well tested implementations ([and results](https://github.com/openai/baselines-results)) of reinforcement learning algorithms from OpenAI
- PyTorch Deep RL - Popular deep RL algorithm implementations with PyTorch
- ChainerRL - Popular deep RL algorithm implementations with Chainer
- Black-DROPS - Modular and generic code for the model-based policy search Black-DROPS algorithm (IROS 2017 paper) and easy integration with the [DART](http://dartsim.github.io/) simulator
- Gold - A reinforcement learning library for Golang.
- Jumanji - A Suite of Industry-Driven Hardware-Accelerated RL Environments written in JAX.
- Pole-Cart Problem
- Q-learning Controller
- BeliefBox - Bayesian reinforcement learning library and toolkit
- Maja - Machine learning framework for problems in Reinforcement Learning in python
-
-
Online Demos
-
Human Computer Interaction
- Deep Q-Learning Demo - A deep Q learning demonstration using ConvNetJS
- Reinforcement Learning Demo - A reinforcement learning demo using reinforcejs by Andrej Karpathy
-
-
Open Source Reinforcement Learning Platforms
-
Human Computer Interaction
- Microsoft AirSim - Open source simulator based on Unreal Engine for autonomous vehicles from Microsoft AI & Research.
- OpenAI gym - A toolkit for developing and comparing reinforcement learning algorithms
- OpenAI universe - A software platform for measuring and training an AI's general intelligence across the world's supply of games, websites and other applications
- DeepMind Lab - A customisable 3D platform for agent-based AI research
- Project Malmo - A platform for Artificial Intelligence experimentation and research built on top of Minecraft by Microsoft
- Retro Learning Environment - An AI platform for reinforcement learning based on video game emulators. Currently supports SNES and Sega Genesis. Compatible with OpenAI gym.
- UETorch - A Torch plugin for Unreal Engine 4 by Facebook
- TorchCraft - Connecting Torch to StarCraft
- garage - A framework for reproducible reinformcement learning research, fully compatible with OpenAI Gym and DeepMind Control Suite (successor to rllab)
- TensorForce - Practical deep reinforcement learning on TensorFlow with Gitter support and OpenAI Gym/Universe/DeepMind Lab integration.
- OpenAI lab - An experimentation system for Reinforcement Learning using OpenAI Gym, Tensorflow, and Keras.
- keras-rl - State-of-the art deep reinforcement learning algorithms in Keras designed for compatibility with OpenAI.
- MAgent - A Platform for Many-agent Reinforcement Learning.
- SLM Lab - A research framework for Deep Reinforcement Learning using Unity, OpenAI Gym, PyTorch, Tensorflow.
- Unity ML Agents - Create reinforcement learning environments using the Unity Editor
- DI-engine - DI-engine is a generalized Decision Intelligence engine. It supports most basic deep reinforcement learning (DRL) algorithms, such as DQN, PPO, SAC, and domain-specific algorithms like QMIX in multi-agent RL, GAIL in inverse RL, and RND in exploration problems.
- Intel Coach - Coach is a python reinforcement learning research framework containing implementation of many state-of-the-art algorithms.
-
-
Theory
-
Books
-
Lectures
- UCL
- Lecture 8: Markov Decision Processes 1
- Lecture 9: Markov Decision Processes 2
- Lecture 10: Reinforcement Learning 1
- Lecture 11: Reinforcement Learning 2
- Stanford
- UC Berkeley
- CMU
- MIT
- Lecture 2: Deep Reinforcement Learning for Motion Planning
- Introduction to AI for video games
- Monte Carlo Prediction
- Q learning explained
- Solving the basic game of Pong
- Actor Critic Algorithms
- War Robots
- Mutual Information
- Reinforcement Learning: A Six Part Series
- The Bellman Equations, Dynamic Programming, and Generalized Policy Iteration
- Monte Carlo And Off-Policy Methods
- TD Learning, Sarsa, and Q-Learning
- UCL
- UCL
- Udacity (Georgia Tech.)
- Introduction to AI for video games
- Monte Carlo Prediction
- Q learning explained
- Solving the basic game of Pong
- Actor Critic Algorithms
- War Robots
- Mutual Information
- Reinforcement Learning: A Six Part Series
- The Bellman Equations, Dynamic Programming, and Generalized Policy Iteration
- Monte Carlo And Off-Policy Methods
- TD Learning, Sarsa, and Q-Learning
- Reinforcement Learning: A Six Part Series
- The Bellman Equations, Dynamic Programming, and Generalized Policy Iteration
- Monte Carlo And Off-Policy Methods
- TD Learning, Sarsa, and Q-Learning
- Stanford
- Lecture 2: Deep Reinforcement Learning for Motion Planning
-
Programming Languages
Categories
Sub Categories
Keywords
reinforcement-learning
18
deep-learning
8
deep-reinforcement-learning
8
machine-learning
7
dqn
5
tensorflow
4
pytorch
4
python
3
actor-critic
3
policy-gradient
3
keras
2
a2c
2
a3c
2
ppo
2
ddpg
2
benchmark
1
categorical-dqn
1
deeprl
1
double-dqn
1
dueling-network-architecture
1
option-critic
1
option-critic-architecture
1
prioritized-experience-replay
1
sac
1
quantile-regression
1
rainbow
1
td3
1
multi-agent
1
course-materials
1
git-course
1
mooc
1
pytorch-tutorials
1
bwapi
1
starcraft
1
torch
1
torchcraft
1
neural-networks
1
unity
1
unity3d
1
reproducibility
1
rl-algorithms
1
artificial-intelligence
1
chainer
1
game-engine
1
machine
1
julia
1
jax
1
research
1
experiment
1
openai
1