{"id":13564851,"url":"https://github.com/jimkon/Deep-Reinforcement-Learning-in-Large-Discrete-Action-Spaces","last_synced_at":"2025-04-03T22:30:26.619Z","repository":{"id":41172765,"uuid":"112010646","full_name":"jimkon/Deep-Reinforcement-Learning-in-Large-Discrete-Action-Spaces","owner":"jimkon","description":"Implementation of the algorithm in Python 3, TensorFlow and OpenAI Gym","archived":false,"fork":false,"pushed_at":"2018-03-01T15:47:02.000Z","size":218433,"stargazers_count":173,"open_issues_count":5,"forks_count":54,"subscribers_count":9,"default_branch":"master","last_synced_at":"2024-11-04T18:45:20.380Z","etag":null,"topics":["ddpg","deep-reinforcement-learning","discrete-actions","wolpertinger"],"latest_commit_sha":null,"homepage":null,"language":"Python","has_issues":true,"has_wiki":null,"has_pages":null,"mirror_url":null,"source_name":null,"license":"mit","status":null,"scm":"git","pull_requests_enabled":true,"icon_url":"https://github.com/jimkon.png","metadata":{"files":{"readme":"README.md","changelog":null,"contributing":null,"funding":null,"license":"LICENSE","code_of_conduct":null,"threat_model":null,"audit":null,"citation":null,"codeowners":null,"security":null,"support":null}},"created_at":"2017-11-25T14:41:04.000Z","updated_at":"2024-11-04T08:37:14.000Z","dependencies_parsed_at":"2022-07-14T09:22:30.964Z","dependency_job_id":null,"html_url":"https://github.com/jimkon/Deep-Reinforcement-Learning-in-Large-Discrete-Action-Spaces","commit_stats":null,"previous_names":[],"tags_count":2,"template":false,"template_full_name":null,"repository_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/jimkon%2FDeep-Reinforcement-Learning-in-Large-Discrete-Action-Spaces","tags_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/jimkon%2FDeep-Reinforcement-Learning-in-Large-Discrete-Action-Spaces/tags","releases_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/jimkon%2FDeep-Reinforcement-Learning-in-Large-Discrete-Action-Spaces/releases","manifests_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/jimkon%2FDeep-Reinforcement-Learning-in-Large-Discrete-Action-Spaces/manifests","owner_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners/jimkon","download_url":"https://codeload.github.com/jimkon/Deep-Reinforcement-Learning-in-Large-Discrete-Action-Spaces/tar.gz/refs/heads/master","host":{"name":"GitHub","url":"https://github.com","kind":"github","repositories_count":247089588,"owners_count":20881797,"icon_url":"https://github.com/github.png","version":null,"created_at":"2022-05-30T11:31:42.601Z","updated_at":"2022-07-04T15:15:14.044Z","host_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub","repositories_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories","repository_names_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repository_names","owners_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners"}},"keywords":["ddpg","deep-reinforcement-learning","discrete-actions","wolpertinger"],"created_at":"2024-08-01T13:01:37.013Z","updated_at":"2025-04-03T22:30:21.584Z","avatar_url":"https://github.com/jimkon.png","language":"Python","funding_links":[],"categories":["Python"],"sub_categories":[],"readme":"# Deep-Reinforcement-Learning-in-Large-Discrete-Action-Spaces\nLink to [paper](https://arxiv.org/abs/1512.07679)\n\nImplementation of the algorithm in Python 3, TensorFlow and OpenAI Gym.\n\n\n\nThis paper introduces Wolpertinger training algorithm that extends the Deep Deterministic Policy Gradient training algorithm introduced in [this](https://arxiv.org/abs/1509.02971) paper.\n\nI used and extended  **stevenpjg**'s implementation of **DDPG** algorithm found [here](https://github.com/stevenpjg/ddpg-aigym) licensed under the MIT license.\n\nMaster is currently **only for continuous action spaces**.\n\nThe branch discrete-and-continuous provides the ability to use the discrete environments of the gym. \n","project_url":"https://awesome.ecosyste.ms/api/v1/projects/github.com%2Fjimkon%2FDeep-Reinforcement-Learning-in-Large-Discrete-Action-Spaces","html_url":"https://awesome.ecosyste.ms/projects/github.com%2Fjimkon%2FDeep-Reinforcement-Learning-in-Large-Discrete-Action-Spaces","lists_url":"https://awesome.ecosyste.ms/api/v1/projects/github.com%2Fjimkon%2FDeep-Reinforcement-Learning-in-Large-Discrete-Action-Spaces/lists"}