Ecosyste.ms: Awesome
An open API service indexing awesome lists of open source software.
https://github.com/chiphuyen/machine-learning-systems-design
A booklet on machine learning systems design with exercises. NOT the repo for the book "Designing Machine Learning Systems"
https://github.com/chiphuyen/machine-learning-systems-design
data-science machine-learning-production mlops
Last synced: about 1 month ago
JSON representation
A booklet on machine learning systems design with exercises. NOT the repo for the book "Designing Machine Learning Systems"
- Host: GitHub
- URL: https://github.com/chiphuyen/machine-learning-systems-design
- Owner: chiphuyen
- Created: 2019-11-24T23:16:28.000Z (about 5 years ago)
- Default Branch: master
- Last Pushed: 2023-04-15T16:33:50.000Z (over 1 year ago)
- Last Synced: 2024-10-23T04:12:21.919Z (3 months ago)
- Topics: data-science, machine-learning-production, mlops
- Language: HTML
- Homepage: https://huyenchip.com/machine-learning-systems-design/toc.html
- Size: 1.34 MB
- Stars: 9,001
- Watchers: 301
- Forks: 1,418
- Open Issues: 10
-
Metadata Files:
- Readme: README.md
Awesome Lists containing this project
- Algorithms-Cheatsheet-Resources - A booklet on machine learning systems design with exercises
README
# Machine Learning Systems Design
**Read this booklet [here](https://huyenchip.com/machine-learning-systems-design/toc.html).**
>>This booklet was my initial attempt to write about machine learning systems design back in 2019. My understanding of the topic has gone through significant iterations since then. My book [Designing Machine Learning Systems](https://www.amazon.com/Designing-Machine-Learning-Systems-Production-Ready/dp/1098107969) (O'Reilly, June 2022) is much more comprehensive and up-to-date. [The new book's repo](https://github.com/chiphuyen/dmls-book) contains the full table of contents, chapter summaries, and random thoughts on MLOps tooling.
This booklet covers four main steps of designing a machine learning system:
1. Project setup
2. Data pipeline
3. Modeling: selecting, training, and debugging
4. Serving: testing, deploying, and maintainingIt comes with links to practical resources that explain each aspect in more details. It also suggests case studies written by machine learning engineers at major tech companies who have deployed machine learning systems to solve real-world problems.
At the end, the booklet contains 27 open-ended machine learning systems design questions that might come up in machine learning interviews. The answers for these questions will be published in the book **Machine Learning Interviews**. You can look at and contribute to community answers to these questions on GitHub [here](https://github.com/chiphuyen/machine-learning-systems-design/tree/master/answers). You can read more about the book and sign up for the book's mailing list [here](https://huyenchip.com/2019/07/21/machine-learning-interviews.html).
## Contribute
This is work-in-progress so any type of contribution is very much appreciated. Here are a few ways you can contribute:1. Improve the text by fixing any lexical, grammatical, or technical error
1. Add more relevant resources to each aspect of the machine learning project flow
1. Add/edit questions
1. Add/edit answers
1. OtherThis book was created using the wonderful [`magicbook`](https://github.com/magicbookproject/magicbook) package. For detailed instructions on how to use the package, see their GitHub repo. The package requires that you have `node`. If you're on Mac, you can install `node` using:
```
brew install node
```Install `magicbook` with:
```
npm install magicbook
```Clone this repository:
```
git clone https://github.com/chiphuyen/machine-learning-systems-design.git
cd machine-learning-systems-design
```After you've made changes to the content in the `content` folder, you can build the booklet by the following steps:
```
magicbook build
```You'll find the generated HTML and PDF files in the folder `build`.
## Acknowledgment
I'd like to thank Ben Krause for being a great friend and helping me with this draft!
## Citation