Ecosyste.ms: Awesome
An open API service indexing awesome lists of open source software.
https://github.com/open-mmlab/mmagic
OpenMMLab Multimodal Advanced, Generative, and Intelligent Creation Toolbox. Unlock the magic 🪄: Generative-AI (AIGC), easy-to-use APIs, awsome model zoo, diffusion models, for text-to-image generation, image/video restoration/enhancement, etc.
https://github.com/open-mmlab/mmagic
aigc computer-vision deep-learning diffusion diffusion-models generative-adversarial-network generative-ai image-editing image-generation image-processing image-synthesis inpainting matting pytorch super-resolution text2image video-frame-interpolation video-interpolation video-super-resolution
Last synced: 5 days ago
JSON representation
OpenMMLab Multimodal Advanced, Generative, and Intelligent Creation Toolbox. Unlock the magic 🪄: Generative-AI (AIGC), easy-to-use APIs, awsome model zoo, diffusion models, for text-to-image generation, image/video restoration/enhancement, etc.
- Host: GitHub
- URL: https://github.com/open-mmlab/mmagic
- Owner: open-mmlab
- License: apache-2.0
- Created: 2019-08-23T13:04:29.000Z (over 5 years ago)
- Default Branch: main
- Last Pushed: 2024-08-06T07:19:40.000Z (5 months ago)
- Last Synced: 2024-10-29T14:50:16.784Z (about 2 months ago)
- Topics: aigc, computer-vision, deep-learning, diffusion, diffusion-models, generative-adversarial-network, generative-ai, image-editing, image-generation, image-processing, image-synthesis, inpainting, matting, pytorch, super-resolution, text2image, video-frame-interpolation, video-interpolation, video-super-resolution
- Language: Jupyter Notebook
- Homepage: https://mmagic.readthedocs.io/en/latest/
- Size: 31.3 MB
- Stars: 6,927
- Watchers: 97
- Forks: 1,057
- Open Issues: 66
-
Metadata Files:
- Readme: README.md
- Contributing: .github/CONTRIBUTING.md
- License: LICENSE
- Code of conduct: .github/CODE_OF_CONDUCT.md
- Citation: CITATION.cff
Awesome Lists containing this project
- AiTreasureBox - open-mmlab/mmagic - 12-20_6994_-1](https://img.shields.io/github/stars/open-mmlab/mmagic.svg) |OpenMMLab Multimodal Advanced, Generative, and Intelligent Creation Toolbox| (Repos)
- StarryDivineSky - open-mmlab/mmagic
README
Multimodal Advanced, Generative, and Intelligent Creation (MMagic [em'mædʒɪk])
[![PyPI](https://badge.fury.io/py/mmagic.svg)](https://pypi.org/project/mmagic/)
[![docs](https://img.shields.io/badge/docs-latest-blue)](https://mmagic.readthedocs.io/en/latest/)
[![badge](https://github.com/open-mmlab/mmagic/workflows/build/badge.svg)](https://github.com/open-mmlab/mmagic/actions)
[![codecov](https://codecov.io/gh/open-mmlab/mmagic/branch/master/graph/badge.svg)](https://codecov.io/gh/open-mmlab/mmagic)
[![license](https://img.shields.io/github/license/open-mmlab/mmagic.svg)](https://github.com/open-mmlab/mmagic/blob/main/LICENSE)
[![open issues](https://isitmaintained.com/badge/open/open-mmlab/mmagic.svg)](https://github.com/open-mmlab/mmagic/issues)
[![issue resolution](https://isitmaintained.com/badge/resolution/open-mmlab/mmagic.svg)](https://github.com/open-mmlab/mmagic/issues)
[![Open in OpenXLab](https://cdn-static.openxlab.org.cn/app-center/openxlab_demo.svg)](https://openxlab.org.cn/apps/detail/%E6%94%BF%E6%9D%B0/OpenMMLab-Projects)[📘Documentation](https://mmagic.readthedocs.io/en/latest/) |
[🛠️Installation](https://mmagic.readthedocs.io/en/latest/get_started/install.html) |
[📊Model Zoo](https://mmagic.readthedocs.io/en/latest/model_zoo/overview.html) |
[🆕Update News](https://mmagic.readthedocs.io/en/latest/changelog.html) |
[🚀Ongoing Projects](https://github.com/open-mmlab/mmagic/projects) |
[🤔Reporting Issues](https://github.com/open-mmlab/mmagic/issues)English | [简体中文](README_zh-CN.md)
### New release [**MMagic v1.2.0**](https://github.com/open-mmlab/mmagic/releases/tag/v1.2.0) \[18/12/2023\]:
- An advanced and powerful inpainting algorithm named PowerPaint is released in our repository. [Click to View](https://github.com/open-mmlab/mmagic/tree/main/projects/powerpaint)
We are excited to announce the release of MMagic v1.0.0 that inherits from [MMEditing](https://github.com/open-mmlab/mmediting) and [MMGeneration](https://github.com/open-mmlab/mmgeneration).
After iterative updates with OpenMMLab 2.0 framework and merged with MMGeneration, MMEditing has become a powerful tool that supports low-level algorithms based on both GAN and CNN. Today, MMEditing embraces Generative AI and transforms into a more advanced and comprehensive AIGC toolkit: **MMagic** (**M**ultimodal **A**dvanced, **G**enerative, and **I**ntelligent **C**reation). MMagic will provide more agile and flexible experimental support for researchers and AIGC enthusiasts, and help you on your AIGC exploration journey.
We highlight the following new features.
**1. New Models**
We support 11 new models in 4 new tasks.
- Text2Image / Diffusion
- ControlNet
- DreamBooth
- Stable Diffusion
- Disco Diffusion
- GLIDE
- Guided Diffusion
- 3D-aware Generation
- EG3D
- Image Restoration
- NAFNet
- Restormer
- SwinIR
- Image Colorization
- InstColorization**2. Magic Diffusion Model**
For the Diffusion Model, we provide the following "magic" :
- Support image generation based on Stable Diffusion and Disco Diffusion.
- Support Finetune methods such as Dreambooth and DreamBooth LoRA.
- Support controllability in text-to-image generation using ControlNet.
- Support acceleration and optimization strategies based on xFormers to improve training and inference efficiency.
- Support video generation based on MultiFrame Render.
- Support calling basic models and sampling strategies through DiffuserWrapper.**3. Upgraded Framework**
By using MMEngine and MMCV of OpenMMLab 2.0 framework, MMagic has upgraded in the following new features:
- Refactor DataSample to support the combination and splitting of batch dimensions.
- Refactor DataPreprocessor and unify the data format for various tasks during training and inference.
- Refactor MultiValLoop and MultiTestLoop, supporting the evaluation of both generation-type metrics (e.g. FID) and reconstruction-type metrics (e.g. SSIM), and supporting the evaluation of multiple datasets at once.
- Support visualization on local files or using tensorboard and wandb.
- Support for 33+ algorithms accelerated by Pytorch 2.0.**MMagic** has supported all the tasks, models, metrics, and losses in [MMEditing](https://github.com/open-mmlab/mmediting) and [MMGeneration](https://github.com/open-mmlab/mmgeneration) and unifies interfaces of all components based on [MMEngine](https://github.com/open-mmlab/mmengine) 😍.
Please refer to [changelog.md](docs/en/changelog.md) for details and release history.
Please refer to [migration documents](docs/en/migration/overview.md) to migrate from [old version](https://github.com/open-mmlab/mmagic/tree/0.x) MMEditing 0.x to new version MMagic 1.x .
## 📄 Table of Contents
- [📖 Introduction](#-introduction)
- [🙌 Contributing](#-contributing)
- [🛠️ Installation](#️-installation)
- [📊 Model Zoo](#-model-zoo)
- [🤝 Acknowledgement](#-acknowledgement)
- [🖊️ Citation](#️-citation)
- [🎫 License](#-license)
- [🏗️ ️OpenMMLab Family](#️-️openmmlab-family)## 📖 Introduction
MMagic (**M**ultimodal **A**dvanced, **G**enerative, and **I**ntelligent **C**reation) is an advanced and comprehensive AIGC toolkit that inherits from [MMEditing](https://github.com/open-mmlab/mmediting) and [MMGeneration](https://github.com/open-mmlab/mmgeneration). It is an open-source image and video editing&generating toolbox based on PyTorch. It is a part of the [OpenMMLab](https://openmmlab.com/) project.
Currently, MMagic support multiple image and video generation/editing tasks.
https://user-images.githubusercontent.com/49083766/233564593-7d3d48ed-e843-4432-b610-35e3d257765c.mp4
### ✨ Major features
- **State of the Art Models**
MMagic provides state-of-the-art generative models to process, edit and synthesize images and videos.
- **Powerful and Popular Applications**
MMagic supports popular and contemporary image restoration, text-to-image, 3D-aware generation, inpainting, matting, super-resolution and generation applications. Specifically, MMagic supports fine-tuning for stable diffusion and many exciting diffusion's application such as ControlNet Animation with SAM. MMagic also supports GAN interpolation, GAN projection, GAN manipulations and many other popular GAN’s applications. It’s time to begin your AIGC exploration journey!
- **Efficient Framework**
By using MMEngine and MMCV of OpenMMLab 2.0 framework, MMagic decompose the editing framework into different modules and one can easily construct a customized editor framework by combining different modules. We can define the training process just like playing with Legos and provide rich components and strategies. In MMagic, you can complete controls on the training process with different levels of APIs. With the support of [MMSeparateDistributedDataParallel](https://github.com/open-mmlab/mmengine/blob/main/mmengine/model/wrappers/seperate_distributed.py), distributed training for dynamic architectures can be easily implemented.
### ✨ Best Practice
- The best practice on our main branch works with **Python 3.9+** and **PyTorch 2.0+**.
## 🙌 Contributing
More and more community contributors are joining us to make our repo better. Some recent projects are contributed by the community including:
- [SDXL](configs/stable_diffusion_xl/README.md) is contributed by @okotaku.
- [AnimateDiff](configs/animatediff/README.md) is contributed by @ElliotQi.
- [ViCo](configs/vico/README.md) is contributed by @FerryHuang.
- [DragGan](configs/draggan/README.md) is contributed by @qsun1.
- [FastComposer](configs/fastcomposer/README.md) is contributed by @xiaomile.[Projects](projects/README.md) is opened to make it easier for everyone to add projects to MMagic.
We appreciate all contributions to improve MMagic. Please refer to [CONTRIBUTING.md](https://github.com/open-mmlab/mmcv/blob/main/CONTRIBUTING.md) in MMCV and [CONTRIBUTING.md](https://github.com/open-mmlab/mmengine/blob/main/CONTRIBUTING.md) in MMEngine for more details about the contributing guideline.
## 🛠️ Installation
MMagic depends on [PyTorch](https://pytorch.org/), [MMEngine](https://github.com/open-mmlab/mmengine) and [MMCV](https://github.com/open-mmlab/mmcv).
Below are quick steps for installation.**Step 1.**
Install PyTorch following [official instructions](https://pytorch.org/get-started/locally/).**Step 2.**
Install MMCV, MMEngine and MMagic with [MIM](https://github.com/open-mmlab/mim).```shell
pip3 install openmim
mim install mmcv>=2.0.0
mim install mmengine
mim install mmagic
```**Step 3.**
Verify MMagic has been successfully installed.```shell
cd ~
python -c "import mmagic; print(mmagic.__version__)"
# Example output: 1.0.0
```**Getting Started**
After installing MMagic successfully, now you are able to play with MMagic! To generate an image from text, you only need several lines of codes by MMagic!
```python
from mmagic.apis import MMagicInferencer
sd_inferencer = MMagicInferencer(model_name='stable_diffusion')
text_prompts = 'A panda is having dinner at KFC'
result_out_dir = 'output/sd_res.png'
sd_inferencer.infer(text=text_prompts, result_out_dir=result_out_dir)
```Please see [quick run](docs/en/get_started/quick_run.md) and [inference](docs/en/user_guides/inference.md) for the basic usage of MMagic.
**Install MMagic from source**
You can also experiment on the latest developed version rather than the stable release by installing MMagic from source with the following commands:
```shell
git clone https://github.com/open-mmlab/mmagic.git
cd mmagic
pip3 install -e .
```Please refer to [installation](docs/en/get_started/install.md) for more detailed instruction.
## 📊 Model Zoo
Supported algorithms
Conditional GANs
Unconditional GANs
Image Restoration
Image Super-Resolution
- DCGAN (ICLR'2016)
- WGAN-GP (NeurIPS'2017)
- LSGAN (ICCV'2017)
- GGAN (ArXiv'2017)
- PGGAN (ICLR'2018)
- SinGAN (ICCV'2019)
- StyleGANV1 (CVPR'2019)
- StyleGANV2 (CVPR'2019)
- StyleGANV3 (NeurIPS'2021)
- DragGan (2023)
- SRCNN (TPAMI'2015)
- SRResNet&SRGAN (CVPR'2016)
- EDSR (CVPR'2017)
- ESRGAN (ECCV'2018)
- RDN (CVPR'2018)
- DIC (CVPR'2020)
- TTSR (CVPR'2020)
- GLEAN (CVPR'2021)
- LIIF (CVPR'2021)
- Real-ESRGAN (ICCVW'2021)
Video Super-Resolution
Video Interpolation
Image Colorization
Image Translation
- EDVR (CVPR'2018)
- TOF (IJCV'2019)
- TDAN (CVPR'2020)
- BasicVSR (CVPR'2021)
- IconVSR (CVPR'2021)
- BasicVSR++ (CVPR'2022)
- RealBasicVSR (CVPR'2022)
Inpainting
Matting
Text-to-Image(Video)
3D-aware Generation
- Global&Local (ToG'2017)
- DeepFillv1 (CVPR'2018)
- PConv (ECCV'2018)
- DeepFillv2 (CVPR'2019)
- AOT-GAN (TVCG'2019)
- Stable Diffusion Inpainting (CVPR'2022)
- GLIDE (NeurIPS'2021)
- Guided Diffusion (NeurIPS'2021)
- Disco-Diffusion (2022)
- Stable-Diffusion (2022)
- DreamBooth (2022)
- Textual Inversion (2022)
- Prompt-to-Prompt (2022)
- Null-text Inversion (2022)
- ControlNet (2023)
- ControlNet Animation (2023)
- Stable Diffusion XL (2023)
- AnimateDiff (2023)
- ViCo (2023)
- FastComposer (2023)
- PowerPaint (2023)
Please refer to [model_zoo](https://mmagic.readthedocs.io/en/latest/model_zoo/overview.html) for more details.
## 🤝 Acknowledgement
MMagic is an open source project that is contributed by researchers and engineers from various colleges and companies. We wish that the toolbox and benchmark could serve the growing research community by providing a flexible toolkit to reimplement existing methods and develop their own new methods.
We appreciate all the contributors who implement their methods or add new features, as well as users who give valuable feedbacks. Thank you all!
## 🖊️ Citation
If MMagic is helpful to your research, please cite it as below.
```bibtex
@misc{mmagic2023,
title = {{MMagic}: {OpenMMLab} Multimodal Advanced, Generative, and Intelligent Creation Toolbox},
author = {{MMagic Contributors}},
howpublished = {\url{https://github.com/open-mmlab/mmagic}},
year = {2023}
}
```
```bibtex
@misc{mmediting2022,
title = {{MMEditing}: {OpenMMLab} Image and Video Editing Toolbox},
author = {{MMEditing Contributors}},
howpublished = {\url{https://github.com/open-mmlab/mmediting}},
year = {2022}
}
```
## 🎫 License
This project is released under the [Apache 2.0 license](LICENSE).
Please refer to [LICENSES](LICENSE) for the careful check, if you are using our code for commercial matters.
## 🏗️ ️OpenMMLab Family
- [MMEngine](https://github.com/open-mmlab/mmengine): OpenMMLab foundational library for training deep learning models.
- [MMCV](https://github.com/open-mmlab/mmcv): OpenMMLab foundational library for computer vision.
- [MIM](https://github.com/open-mmlab/mim): MIM installs OpenMMLab packages.
- [MMPreTrain](https://github.com/open-mmlab/mmpretrain): OpenMMLab Pre-training Toolbox and Benchmark.
- [MMDetection](https://github.com/open-mmlab/mmdetection): OpenMMLab detection toolbox and benchmark.
- [MMDetection3D](https://github.com/open-mmlab/mmdetection3d): OpenMMLab's next-generation platform for general 3D object detection.
- [MMRotate](https://github.com/open-mmlab/mmrotate): OpenMMLab rotated object detection toolbox and benchmark.
- [MMSegmentation](https://github.com/open-mmlab/mmsegmentation): OpenMMLab semantic segmentation toolbox and benchmark.
- [MMOCR](https://github.com/open-mmlab/mmocr): OpenMMLab text detection, recognition, and understanding toolbox.
- [MMPose](https://github.com/open-mmlab/mmpose): OpenMMLab pose estimation toolbox and benchmark.
- [MMHuman3D](https://github.com/open-mmlab/mmhuman3d): OpenMMLab 3D human parametric model toolbox and benchmark.
- [MMSelfSup](https://github.com/open-mmlab/mmselfsup): OpenMMLab self-supervised learning toolbox and benchmark.
- [MMRazor](https://github.com/open-mmlab/mmrazor): OpenMMLab model compression toolbox and benchmark.
- [MMFewShot](https://github.com/open-mmlab/mmfewshot): OpenMMLab fewshot learning toolbox and benchmark.
- [MMAction2](https://github.com/open-mmlab/mmaction2): OpenMMLab's next-generation action understanding toolbox and benchmark.
- [MMTracking](https://github.com/open-mmlab/mmtracking): OpenMMLab video perception toolbox and benchmark.
- [MMFlow](https://github.com/open-mmlab/mmflow): OpenMMLab optical flow toolbox and benchmark.
- [MMagic](https://github.com/open-mmlab/mmagic): OpenMMLab Multimodal Advanced, Generative, and Intelligent Creation Toolbox.
- [MMDeploy](https://github.com/open-mmlab/mmdeploy): OpenMMLab model deployment framework.