awesome-deep-model-compression
Awesome Deep Model Compression
https://github.com/chadHGY/awesome-deep-model-compression
Last synced: 16 days ago
JSON representation
-
Articles
-
Blogs
- Pruning deep neural networks to make them fast and small - 16 based Dogs-vs-Cats classifier is made x3 faster and x4 smaller.
- All The Ways You Can Compress BERT - An overview of different compression methods for large NLP models (BERT) based on different characteristics and compares their results.
- Deep Learning Model Compression
- Do We Really Need Model Compression
- All The Ways You Can Compress BERT - An overview of different compression methods for large NLP models (BERT) based on different characteristics and compares their results.
- Do We Really Need Model Compression
-
-
Papers
-
Pruning
- Paper - research/lottery-ticket-hypothesis)
-
-
Tools
-
Cross Platform
-
Hard-ware Integration
- TensorRT (NVIDIA)
- How to Convert a Model from PyTorch to TensorRT and Speed Up Inference
- Introduction
- torch2trt - AI-IOT/torch2trt.svg?style=social)](https://github.com/NVIDIA-AI-IOT/torch2trt)
- TVM (Apache) - glr-tutorials-frontend-from-tensorflow-py)] [](https://github.com/apache/tvm)
- Pytorch Glow
- CoreML (Apple)
- ~~Intel Nervana Neon~~
- Tensorflow Lite (Google)
-
Libraries
- torch.nn.utils.prune
- Neural Network Intelligence
- ![Star on GitHub
- ![Star on GitHub
- paper
- ![Star on GitHub - Pruning)
- ![Star on GitHub
- ![Star on GitHub - marple-dev/model_compression)
- Reposhub
- ![Star on GitHub - optimization)
- TensorFlow Model Optimization Toolkit — Pruning API
- ![Star on GitHub
- Condensa
- IntelLabs distiller
- Documentation
- Torch-Pruning
- CompressAI
- Model Compression
- TensorFlow Model Optimization Toolkit
- XNNPACK
- paper
-
Programming Languages
Categories
Sub Categories
Keywords
pytorch
6
pruning
5
deep-learning
4
machine-learning
4
deep-neural-networks
4
quantization
4
neural-network
3
model-compression
3
tensorflow
2
compression
2
truncated-svd
2
regularization
2
pruning-structures
2
inference
2
onnx
2
network-compression
2
jupyter-notebook
2
group-lasso
2
early-exit
2
distillation
2
automl-for-compression
2
python
2
performance
2
optimization
1
ml
1
fast
1
keras
1
mkl
1
neon
1
structured-pruning
1
structural-pruning
1
network-pruning
1
efficient-deep-learning
1
depgraph
1
cvpr2023
1
channel-pruning
1
vulkan
1
classification
1
jetson-nano
1
jetson-tx2
1
jetson-xavier
1
tensorrt
1
model-pruning
1
lottey-ticket-hypothesis
1
convolutional-neural-network
1
convolutional-neural-networks
1
cpu
1
inference-optimization
1
matrix-multiplication
1
mobile-inference
1