{"id":19622299,"url":"https://github.com/camigord/drl_papernotes","last_synced_at":"2026-08-05T23:30:20.825Z","repository":{"id":253787517,"uuid":"79914306","full_name":"camigord/DRL_papernotes","owner":"camigord","description":"Notes and comments about Deep Reinforcement Learning papers","archived":false,"fork":false,"pushed_at":"2018-01-04T13:02:52.000Z","size":2505,"stargazers_count":77,"open_issues_count":0,"forks_count":15,"subscribers_count":13,"default_branch":"master","last_synced_at":"2025-02-26T19:24:45.045Z","etag":null,"topics":["deep-reinforcement-learning","hierarchical-reinforcement-learning","intrinsic-motivation","papers","reinforcement-learning"],"latest_commit_sha":null,"homepage":null,"language":null,"has_issues":true,"has_wiki":null,"has_pages":null,"mirror_url":null,"source_name":null,"license":null,"status":null,"scm":"git","pull_requests_enabled":true,"icon_url":"https://github.com/camigord.png","metadata":{"files":{"readme":"README.md","changelog":null,"contributing":null,"funding":null,"license":null,"code_of_conduct":null,"threat_model":null,"audit":null,"citation":null,"codeowners":null,"security":null,"support":null,"governance":null,"roadmap":null,"authors":null,"dei":null,"publiccode":null,"codemeta":null}},"created_at":"2017-01-24T13:32:20.000Z","updated_at":"2024-12-02T08:07:11.000Z","dependencies_parsed_at":"2024-08-21T19:17:16.607Z","dependency_job_id":null,"html_url":"https://github.com/camigord/DRL_papernotes","commit_stats":null,"previous_names":["camigord/drl_papernotes"],"tags_count":0,"template":false,"template_full_name":null,"purl":"pkg:github/camigord/DRL_papernotes","repository_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/camigord%2FDRL_papernotes","tags_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/camigord%2FDRL_papernotes/tags","releases_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/camigord%2FDRL_papernotes/releases","manifests_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/camigord%2FDRL_papernotes/manifests","owner_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners/camigord","download_url":"https://codeload.github.com/camigord/DRL_papernotes/tar.gz/refs/heads/master","sbom_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/camigord%2FDRL_papernotes/sbom","scorecard":null,"host":{"name":"GitHub","url":"https://github.com","kind":"github","repositories_count":286080680,"owners_count":36323916,"icon_url":"https://github.com/github.png","version":null,"created_at":"2022-05-30T11:31:42.601Z","updated_at":"2026-07-20T02:08:10.276Z","status":"online","status_checked_at":"2026-08-05T02:00:06.619Z","response_time":104,"last_error":null,"robots_txt_status":"success","robots_txt_updated_at":"2025-07-24T06:49:26.215Z","robots_txt_url":"https://github.com/robots.txt","online":true,"can_crawl_api":true,"host_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub","repositories_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories","repository_names_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repository_names","owners_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners"}},"keywords":["deep-reinforcement-learning","hierarchical-reinforcement-learning","intrinsic-motivation","papers","reinforcement-learning"],"created_at":"2024-11-11T11:27:06.339Z","updated_at":"2026-08-05T23:30:20.387Z","avatar_url":"https://github.com/camigord.png","language":null,"funding_links":[],"categories":[],"sub_categories":[],"readme":"# Deep Reinforcement Learning papernotes\n\n\u003e _New Hierarchical-Learning section [here](./notes/hierarchical-learning/)._\n\n\n#### 2017-05\n- [Curiosity-driven Exploration by Self-supervised Prediction](./notes/Intrinsic%20Motivation/Curiosity-driven%20Exploration%20by%20self-supervised%20prediction.md) [[arXiv](https://arxiv.org/abs/1705.05363)]\n\n#### 2017-03\n- [Surprised-Based Intrinsic Motivation for Deep Reinforcement Learning](./notes/Intrinsic%20Motivation/Surprised-based%20intrinsic%20motivation%20for%20DRL.md)[[arXiv](https://arxiv.org/abs/1703.01732)]\n- [Virtual-to-real Deep Reinforcement Learning: Continuous Control of Mobile Robots for Mapless Navigation](https://github.com/camigord/DRL_papernotes/blob/master/notes/Virtual-to-real%20%20Deep%20%20Reinforcement%20%20Learning.md) [[arXiv](https://arxiv.org/abs/1703.00420)]\n\n#### 2017-02\n\n- [Neural Map : Structured Memory for Deep Reinforcement Learning](https://github.com/camigord/DRL_papernotes/blob/master/notes/NeuralMapDRL.md) [[arXiv](https://arxiv.org/abs/1702.08360)]\n\n#### 2017-01\n\n- Deep Recurrent Q-Learning for Partially Observable MDPs [[arXiv](https://arxiv.org/abs/1507.06527)]\n\n#### 2016-12\n\n- [Playing Doom with SLAM-Augmented Deep Reinforcement Learning](https://github.com/camigord/DRL_papernotes/blob/master/notes/Playing%20Doom%20with%20SLAM-Augmented%20Deep%20Reinforcement%20Learning.md) [[arXiv](https://arxiv.org/abs/1612.00380)]\n- Learning to predict where to look in interactive environments using deep recurrent q-learning [[arXiv](https://arxiv.org/abs/1612.05753)]\n\n#### 2016-11\n\n- [Deep Reinforcement Learning for Robotic Manipulation with Asynchronous Off-Policy Updates](https://github.com/camigord/DRL_papernotes/blob/master/notes/Deep%20Reinforcement%20Learning%20for%20Robotic%20Manipulation%20with%20Asynchronous.md)[[arXiv](https://arxiv.org/abs/1610.00633)]\n- Reinforcement Learning with Unsupervised Auxiliary Tasks [[arXiv](https://arxiv.org/abs/1611.05397)]\n- [Learning to Navigate in Complex Environments](https://github.com/camigord/DRL_papernotes/blob/master/notes/Learning%20to%20Navigate%20in%20Complex%20Environments.md) [[arXiv](https://arxiv.org/abs/1611.03673)]\n- Learning to reinforcement learn [[arXiv](https://arxiv.org/abs/1611.05763)]\n\n#### 2016-10\n\n- [Hybrid computing using a neural network with dynamic external memory](https://github.com/camigord/DRL_papernotes/blob/master/notes/Hybrid%20computing%20using%20NN%20with%20dynamic%20external%20memory.md) [[nature](http://www.nature.com/nature/journal/v538/n7626/abs/nature20101.html)]\n- A Deep Hierarchical Approach to Lifelong Learning in Minecraft [[arXiv](https://arxiv.org/abs/1604.07255)]\n- Towards Cognitive Exploration through Deep Reinforcement Learning for Mobile Robots [[arXiv](https://arxiv.org/abs/1610.01733)]\n\n#### 2016-09\n\n- [Target-driven Visual Navigation in Indoor Scenes using Deep Reinforcement Learning](https://github.com/camigord/DRL_papernotes/blob/master/notes/Target-driven%20Visual%20Navigation%20in%20Indoor%20Scenes%20using%20Deep%20Reinforcement%20Learning.md) [[arXiv](https://arxiv.org/abs/1609.05143)]\n- Playing FPS Games with Deep Reinforcement Learning [[arXiv](https://arxiv.org/abs/1609.05521)]\n\n#### 2016-08\n\n- [Learning Hand-Eye Coordination for Robotic Grasping with Deep Learning and Large-Scale Data Collection](https://github.com/camigord/DRL_papernotes/blob/master/notes/Learning%20Hand-Eye%20Coordination%20for%20Robotic%20Grasping%20with%20Deep%20Learning%20and%20Large-Scale%20Data%20Collection.md) [[arXiv](https://arxiv.org/abs/1603.02199)]\n\n#### 2016-05\n\n- [Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation](https://github.com/camigord/DRL_papernotes/blob/master/notes/Hierarchical%20Deep%20Reinforcement%20Learning.md) [[arXiv](https://arxiv.org/abs/1604.06057)]\n- Value Iteration Networks [[arXiv](https://arxiv.org/abs/1602.02867)]\n\n#### 2016-04\n\n- End-to-End Training of Deep Visuomotor Policies [[arXiv](https://arxiv.org/abs/1504.00702)]\n\n#### 2016-02\n\n- [Prioritized Experience Replay](https://github.com/camigord/DRL_papernotes/blob/master/notes/Prioritized%20Experience%20Replay.md) [[arXiv](https://arxiv.org/abs/1511.05952)]\n- [Asynchronous Methods for Deep Reinforcement Learning](https://github.com/camigord/DRL_papernotes/blob/master/notes/Asynchronous%20Methods%20for%20Deep%20Reinforcement%20Learning.md) [[arXiv](https://arxiv.org/abs/1602.01783)]\n- Continuous control with deep reinforcement learning [[arXiv](https://arxiv.org/abs/1509.02971)]\n- Graying the black box: Understanding DQNs [[arXiv](https://arxiv.org/abs/1602.02658)]\n\n#### 2015-12\n\n- [Deep Reinforcement Learning with Double Q-learning](https://github.com/camigord/DRL_papernotes/blob/master/notes/Deep%20Reinforcement%20Learning%20with%20Double%20Q-learning.md) [[arXiv](https://arxiv.org/abs/1509.06461)]\n- Deep Attention Recurrent Q-Network [[arXiv](https://arxiv.org/abs/1512.01693)]\n\n#### 2015-11\n\n- [Towards Vision-Based Deep Reinforcement Learning for Robotic Motion Control](https://github.com/camigord/DRL_papernotes/blob/master/notes/Towards%20Vision-Based%20Deep%20Reinforcement%20Learning%20for%20Robotic%20Motion%20Control.md) [[arXiv](https://arxiv.org/abs/1511.03791)]\n\n#### 2015-02\n\n- [Human-level control through deep reinforcement learning](https://github.com/camigord/DRL_papernotes/blob/master/notes/Human-level%20control%20through%20deep%20reinforcement%20learning.md) [[nature](http://www.nature.com/nature/journal/v518/n7540/full/nature14236.html)]\n","project_url":"https://awesome.ecosyste.ms/api/v1/projects/github.com%2Fcamigord%2Fdrl_papernotes","html_url":"https://awesome.ecosyste.ms/projects/github.com%2Fcamigord%2Fdrl_papernotes","lists_url":"https://awesome.ecosyste.ms/api/v1/projects/github.com%2Fcamigord%2Fdrl_papernotes/lists"}