{"id":1166,"slug":"reinforcement-learning","name":"Reinforcement learning","short_description":"Reinforcement learning is a machine learning paradigm where agents learn optimal behavior through environment interaction.","url":"https://github.com/topics/reinforcement-learning","github_count":21936,"created_by":null,"logo_url":null,"released":null,"wikipedia_url":"https://en.wikipedia.org/wiki/Reinforcement_learning","related_topics":[],"aliases":[],"github_url":null,"content":"\u003cp\u003eReinforcement learning is a machine learning paradigm focused on sequential decision-making, in which an autonomous agent learns optimal behavior by interacting with a dynamic environment to maximize cumulative reward signals.\u003c/p\u003e\n","created_at":"2026-03-11T00:18:38.116Z","updated_at":"2026-10-06T00:26:26.413Z","topic_url":"https://awesome.ecosyste.ms/api/v1/topics/reinforcement-learning","html_url":"https://awesome.ecosyste.ms/topics/reinforcement-learning","projects_url":"https://awesome.ecosyste.ms/api/v1/projects?keyword=reinforcement-learning","lists_url":"https://awesome.ecosyste.ms/api/v1/lists?topic=reinforcement-learning"}