{"id":26201686,"url":"https://github.com/naveen88112/clustering_customer_invoice_data","last_synced_at":"2026-04-13T10:32:35.100Z","repository":{"id":281817334,"uuid":"946489472","full_name":"Naveen88112/Clustering_Customer_Invoice_Data","owner":"Naveen88112","description":"Customer Invoice Data Clustering This project uses clustering methods on customer invoice data for segmentation analysis. It preprocesses data, normalizes features, and uses K-Means and DBSCAN to cluster customers according to spending habits and shared locations.","archived":false,"fork":false,"pushed_at":"2025-03-11T09:12:09.000Z","size":42,"stargazers_count":0,"open_issues_count":0,"forks_count":0,"subscribers_count":1,"default_branch":"main","last_synced_at":"2025-03-11T10:27:29.057Z","etag":null,"topics":["clustering","data-preprocessing","data-visualization","numpy","pandas","python","silhouette-score","standardization"],"latest_commit_sha":null,"homepage":"","language":"Jupyter Notebook","has_issues":true,"has_wiki":null,"has_pages":null,"mirror_url":null,"source_name":null,"license":null,"status":null,"scm":"git","pull_requests_enabled":true,"icon_url":"https://github.com/Naveen88112.png","metadata":{"files":{"readme":"README.md","changelog":null,"contributing":null,"funding":null,"license":null,"code_of_conduct":null,"threat_model":null,"audit":null,"citation":null,"codeowners":null,"security":null,"support":null,"governance":null,"roadmap":null,"authors":null,"dei":null,"publiccode":null,"codemeta":null}},"created_at":"2025-03-11T08:13:09.000Z","updated_at":"2025-03-11T09:15:51.000Z","dependencies_parsed_at":"2025-03-11T10:27:30.633Z","dependency_job_id":"a88268bd-9a8a-46dc-a30a-7d1d5e0fae16","html_url":"https://github.com/Naveen88112/Clustering_Customer_Invoice_Data","commit_stats":null,"previous_names":["naveen88112/clustering_customer_invoice_data"],"tags_count":0,"template":false,"template_full_name":null,"repository_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Naveen88112%2FClustering_Customer_Invoice_Data","tags_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Naveen88112%2FClustering_Customer_Invoice_Data/tags","releases_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Naveen88112%2FClustering_Customer_Invoice_Data/releases","manifests_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Naveen88112%2FClustering_Customer_Invoice_Data/manifests","owner_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners/Naveen88112","download_url":"https://codeload.github.com/Naveen88112/Clustering_Customer_Invoice_Data/tar.gz/refs/heads/main","host":{"name":"GitHub","url":"https://github.com","kind":"github","repositories_count":243148059,"owners_count":20243910,"icon_url":"https://github.com/github.png","version":null,"created_at":"2022-05-30T11:31:42.601Z","updated_at":"2022-07-04T15:15:14.044Z","host_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub","repositories_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories","repository_names_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repository_names","owners_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners"}},"keywords":["clustering","data-preprocessing","data-visualization","numpy","pandas","python","silhouette-score","standardization"],"created_at":"2025-03-12T03:23:16.684Z","updated_at":"2025-12-24T10:40:49.064Z","avatar_url":"https://github.com/Naveen88112.png","language":"Jupyter Notebook","funding_links":[],"categories":[],"sub_categories":[],"readme":"Customer Invoice Data Clustering\n\nOverview\nThis project focuses on clustering customer invoice data using machine learning techniques. It aims to segment customers based on their spending behavior and transaction patterns to derive meaningful business insights.\n\nFeatures\n- Data Preprocessing: Standardization of invoice-related numerical features.\n- Clustering Algorithms: K-Means and DBSCAN applied for segmentation.\n- Performance Metrics: Silhouette Score used to evaluate clustering effectiveness.\n- Visualization: Cluster insights represented using Matplotlib.\n\nTechnologies Used\n- Python\n- Pandas \u0026 NumPy\n- Scikit-learn\n- Matplotlib\n\nHow to Run\n1. Clone the repository:\n   \n   \"git clone https://github.com/yourusername/customer-invoice-clustering.git\"\n   \n2. Open the Jupyter Notebook or Google Colab.\n3. Upload the dataset (if required) and execute the cells step by step.\n\nResults \u0026 Insights\n- Customers were segmented based on invoice amounts and transaction patterns.\n- K-Means and DBSCAN clustering methods were compared for effectiveness.\n- Visualization helped understand customer behavior in different segments.\n","project_url":"https://awesome.ecosyste.ms/api/v1/projects/github.com%2Fnaveen88112%2Fclustering_customer_invoice_data","html_url":"https://awesome.ecosyste.ms/projects/github.com%2Fnaveen88112%2Fclustering_customer_invoice_data","lists_url":"https://awesome.ecosyste.ms/api/v1/projects/github.com%2Fnaveen88112%2Fclustering_customer_invoice_data/lists"}