{"id":22045793,"url":"https://github.com/abduldevhub/reddit-data-scrapping","last_synced_at":"2026-05-07T00:39:45.204Z","repository":{"id":231061878,"uuid":"780783426","full_name":"AbdulDevHub/Reddit-Data-Scrapping","owner":"AbdulDevHub","description":"This project is a portfolio of my work on data scraping from Reddit using Python. Read the README file to learn more.","archived":false,"fork":false,"pushed_at":"2024-09-07T21:28:59.000Z","size":8327,"stargazers_count":0,"open_issues_count":0,"forks_count":0,"subscribers_count":1,"default_branch":"main","last_synced_at":"2025-06-03T07:52:21.252Z","etag":null,"topics":["gensim","jupyter-notebook","matplotlib","nltk","pandas","praw","python"],"latest_commit_sha":null,"homepage":"","language":"Jupyter Notebook","has_issues":true,"has_wiki":null,"has_pages":null,"mirror_url":null,"source_name":null,"license":null,"status":null,"scm":"git","pull_requests_enabled":true,"icon_url":"https://github.com/AbdulDevHub.png","metadata":{"files":{"readme":"README.md","changelog":null,"contributing":null,"funding":null,"license":null,"code_of_conduct":null,"threat_model":null,"audit":null,"citation":null,"codeowners":null,"security":null,"support":null,"governance":null,"roadmap":null,"authors":null,"dei":null,"publiccode":null,"codemeta":null}},"created_at":"2024-04-02T06:40:06.000Z","updated_at":"2024-09-07T21:29:03.000Z","dependencies_parsed_at":"2024-05-06T00:40:42.924Z","dependency_job_id":null,"html_url":"https://github.com/AbdulDevHub/Reddit-Data-Scrapping","commit_stats":null,"previous_names":["abduldevhub/reddit-data-scrapping-with-python","abduldevhub/reddit-data-scrapping"],"tags_count":0,"template":false,"template_full_name":null,"purl":"pkg:github/AbdulDevHub/Reddit-Data-Scrapping","repository_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/AbdulDevHub%2FReddit-Data-Scrapping","tags_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/AbdulDevHub%2FReddit-Data-Scrapping/tags","releases_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/AbdulDevHub%2FReddit-Data-Scrapping/releases","manifests_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/AbdulDevHub%2FReddit-Data-Scrapping/manifests","owner_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners/AbdulDevHub","download_url":"https://codeload.github.com/AbdulDevHub/Reddit-Data-Scrapping/tar.gz/refs/heads/main","sbom_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/AbdulDevHub%2FReddit-Data-Scrapping/sbom","host":{"name":"GitHub","url":"https://github.com","kind":"github","repositories_count":267780023,"owners_count":24143201,"icon_url":"https://github.com/github.png","version":null,"created_at":"2022-05-30T11:31:42.601Z","updated_at":"2022-07-04T15:15:14.044Z","status":"online","status_checked_at":"2025-07-29T02:00:12.549Z","response_time":2574,"last_error":null,"robots_txt_status":"success","robots_txt_updated_at":"2025-07-24T06:49:26.215Z","robots_txt_url":"https://github.com/robots.txt","online":true,"can_crawl_api":true,"host_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub","repositories_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories","repository_names_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repository_names","owners_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners"}},"keywords":["gensim","jupyter-notebook","matplotlib","nltk","pandas","praw","python"],"created_at":"2024-11-30T13:15:04.325Z","updated_at":"2026-05-07T00:39:45.176Z","avatar_url":"https://github.com/AbdulDevHub.png","language":"Jupyter Notebook","funding_links":[],"categories":[],"sub_categories":[],"readme":"# Reddit Data Scrapping With Python\n\nWelcome to my repository! This project is a portfolio of my work on data scraping from Reddit using Python. \n\n\u003cimg height=\"400\" src=\"./Banner.png\"\u003e\n\n## Project Overview\n\nThis project involves scraping data from Reddit using the Python Reddit API Wrapper (PRAW), performing data analysis, and visualizing the results. The repository contains 2 Jupyter notebooks with all the utilized code, a PDF and PowerPoint of the findings, and a folder with the produced images/figures.\n\n## Dependencies\n\nThis project uses the following Python libraries:\n\n- `praw`: For interacting with the Reddit API.\n- `pandas`: For data manipulation and analysis.\n- `datetime`: For working with dates and times.\n- `nltk`: For natural language processing tasks.\n- `gensim`: For topic modelling and document similarity analysis.\n- `matplotlib`: For creating static, animated, and interactive visualizations in Python.\n- `pyLDAvis`: For interactive topic model visualization.\n- `wordcloud`: For creating word cloud images.\n\n## Getting Started\n\nTo run the notebook, you can either extract the code to a separate Python file, or just run these two files directly on Jupyter Notebook.\n\nYou'll also need to set up PRAW with your own Reddit API credentials. You can do this in the notebooks where the PRAW instance is initialized.\n\n## Contact\n\nIf you have any questions or feedback, feel free to open an issue or submit a pull request.\n\nEnjoy exploring the repository!\n\n\u003cbr\u003e\n","project_url":"https://awesome.ecosyste.ms/api/v1/projects/github.com%2Fabduldevhub%2Freddit-data-scrapping","html_url":"https://awesome.ecosyste.ms/projects/github.com%2Fabduldevhub%2Freddit-data-scrapping","lists_url":"https://awesome.ecosyste.ms/api/v1/projects/github.com%2Fabduldevhub%2Freddit-data-scrapping/lists"}