{"id":14698,"url":"https://github.com/Franceshe/awesome-generative-models","name":"awesome-generative-models","description":"A collection of awesome generative model papers, frameworks, libraries, software and resources for text, image, video, animation, code generation","projects_count":81,"last_synced_at":"2026-08-08T12:00:38.838Z","repository":{"id":41149548,"uuid":"246799993","full_name":"Franceshe/awesome-generative-models","owner":"Franceshe","description":"A collection of awesome generative model papers, frameworks, libraries, software and resources for text, image, video, animation, code generation","archived":false,"fork":false,"pushed_at":"2021-05-27T14:45:46.000Z","size":64,"stargazers_count":25,"open_issues_count":0,"forks_count":2,"subscribers_count":1,"default_branch":"master","last_synced_at":"2026-07-20T05:04:13.438Z","etag":null,"topics":["biggan-models","code-generation","deep-learning","generative-adversarial-network","generative-art","generative-model","image-generation","image-processing","machine-learning","music-generation","question-answering","question-generation","static-analysis","text-generation","video-generation"],"latest_commit_sha":null,"homepage":"","language":null,"has_issues":true,"has_wiki":null,"has_pages":null,"mirror_url":null,"source_name":null,"license":null,"status":null,"scm":"git","pull_requests_enabled":true,"icon_url":"https://github.com/Franceshe.png","metadata":{"files":{"readme":"README.md","changelog":null,"contributing":null,"funding":null,"license":null,"code_of_conduct":null,"threat_model":null,"audit":null,"citation":null,"codeowners":null,"security":null,"support":null}},"created_at":"2020-03-12T09:56:58.000Z","updated_at":"2025-12-09T13:36:46.000Z","dependencies_parsed_at":"2022-08-25T17:23:59.729Z","dependency_job_id":null,"html_url":"https://github.com/Franceshe/awesome-generative-models","commit_stats":null,"previous_names":[],"tags_count":0,"template":false,"template_full_name":null,"purl":"pkg:github/Franceshe/awesome-generative-models","repository_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Franceshe%2Fawesome-generative-models","tags_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Franceshe%2Fawesome-generative-models/tags","releases_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Franceshe%2Fawesome-generative-models/releases","manifests_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Franceshe%2Fawesome-generative-models/manifests","owner_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners/Franceshe","download_url":"https://codeload.github.com/Franceshe/awesome-generative-models/tar.gz/refs/heads/master","sbom_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Franceshe%2Fawesome-generative-models/sbom","scorecard":null,"host":{"name":"GitHub","url":"https://github.com","kind":"github","repositories_count":286080680,"owners_count":36406705,"icon_url":"https://github.com/github.png","version":null,"created_at":"2022-05-30T11:31:42.601Z","updated_at":"2026-08-06T04:43:03.162Z","status":"online","status_checked_at":"2026-08-08T02:00:08.763Z","response_time":96,"last_error":null,"robots_txt_status":"success","robots_txt_updated_at":"2025-07-24T06:49:26.215Z","robots_txt_url":"https://github.com/robots.txt","online":true,"can_crawl_api":true,"host_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub","repositories_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories","repository_names_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repository_names","owners_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners"}},"created_at":"2024-01-12T20:23:58.048Z","updated_at":"2026-08-08T12:00:38.839Z","primary_language":null,"list_of_lists":false,"displayable":true,"categories":["Resources on related course:","Audio and music generation/processing","Code generation: ML algorithm + Static Analysis","Image Style Transfer","NLP: Text-Sentiment-Analysis:","Text-to-Video:","Image Captioning","Text-to-Animation","Text Generation/NLP","Video Synthesis and Generation:","ProceduralGenerationForGaming","Animation","Computer Vision application in gaming","Cloud computing set up","Full Stack Deep Learning tools in data processing pipeline:","Resources for dev tool:","MEET mind-linked people in forum:","Application:","Design","Reference:","Code to GUI: algorithm to user interface:","Audio Systhesis","Text-to-Image:","Image Synthesis"],"sub_categories":["Map generation","Project:","Other","Text Corpus/dataset","Music-Video Synthesis","Generating voices of various characters: Ideally podcast and gaming etc:","FILM MAKING:","Question Generation:"],"readme":"# awesome-generative-models\nA curated list of awesome generative model frameworks, libraries, software and resources for media production\n\ninspired by [awesome-python](https://github.com/vinta/awesome-python)\n\u003cdiv align=\"center\"\u003e\n\t\u003cimg width=\"500\" height=\"350\" src=\"media/somercloud_1_logo.PNG\" alt=\"Awesome\"\u003e\n\t\u003cbr\u003e\n\t\u003cbr\u003e\n\t\u003chr\u003e\n\t\u003cp\u003e\n\t\t\u003cp\u003e\n\t\t\t\u003csup\u003e\n\t\t\t\t\u003ca href=\"https://github.com/SomerCloud-Studio\"\u003eMy open source work is supported by the community\u003c/a\u003e\n\t\t\t\u003c/sup\u003e\n\u003c/div\u003e\n\n- [Awesome Generative Mdels](#awesome-generative-models)\n\t- [Code generation: ML algorithm + Static Analysis](#CodeGeration)\n    - [Code to GUI: algorithm to user interface](#CodeToGUI)\n\t- [Image Synthesis](#imageSynthesis)\n\t\t- [Image Style Transfer](#ImageStyleTransfer)\n    - [Text Generation/NLP](#NLP)\n\t\t- [Audio and music generation/processing](#AudioAndMusicGenerationProcessing)\n\t\t- [Music-Video Synthesis](#MusicVideoSynthesis)\n    - [Video-Synth](#videoSynth)\n\t- [Procedural Generation for gaming](#ProceduralGenerationForGaming)\n\t- [Image-Generation](#imageGeneration)\n    - [Forum](#forum)\n\n## Code to GUI: algorithm to user interface:\n* [Screenshoot to Code](https://github.com/emilwallner/Screenshot-to-code)\n* [pix2code: Generating Code from a Graphical User Interface Screenshot](https://github.com/tonybeltramelli/pix2code)\n* [Bred Victor: Machine learning in Engineering inference from enviroment](http://worrydream.com/#!/MagicInk)\n\n## Code generation: ML algorithm + Static Analysis\n* [Tabnine: Autocompletion with deep learning](https://www.tabnine.com/blog/deep/)\n* [Kite: Algorithm assisted coder completion ](https://kite.com/)\n* Insignt:\n  Standard code completion tools often still use alphabetical sorting, while Kite uses ML algorithms to infer what a developer is likely trying to do\n## Image Synthesis\n* [SPADE by NVlabs: Synthesizing photorealistic images given an input semantic layout](https://nvlabs.github.io/SPADE/).\n  [code](https://github.com/nvlabs/spade/)\n\n## Image Style Transfer\n* [Artbreeder based on BigGAN models](https://artbreeder.com/)\n   [opensource version](https://github.com/joel-simon/ganbreeder)\n   [BigGAN models](https://tfhub.dev/deepmind/biggan-512/2)\n   [About](https://artbreeder.com/about)\n* A cool artistic project: \n... DGSpitzer(Eddie Hu) a indie game maker Use [Artbreeder based on BigGAN models](https://artbreeder.com/), StyleGAN-Artm Realistic-Neural_Talking_Head_Modelsm First-Order_Modelsm DAIN\nand Topaz Lab for some great work on digital video repair, which colorize black and while video.\n\n* A cool project: [Visual Novel using ML Style transfer](https://medium.com/@pbriod/how-i-used-artificial-intelligence-to-make-my-first-graphic-novel-about-my-trip-to-india-although-ff4af46fd6a3)\n## NLP: Text-Sentiment-Analysis:\n* [DeepMoji:a sentiment analysis model](https://arxiv.org/pdf/1708.00524.pdf) **This model also used in [15.ai](https://15.ai/about), a text-to-speech tool for generating voices of various characters.**\n\n## Text-to-Image:\n* dataset: [Google’s “Quick Draw” open source dataset](https://opensource.google/projects/quickdrawdataset)\n  [github](https://github.com/googlecreativelab/quickdraw-dataset)\n\n## Text-to-Video:\n* AllenNLP: [Imagine This! Scripts to Compositions to Videos](https://prior.allenai.org/projects/craft)\n\n## Image Captioning\n* [Tensorflow core: Image captioning with visual attention](https://www.tensorflow.org/tutorials/text/image_captioning)\n\n## Text-to-Animation\n* [Generating Animations from Screenplays](https://arxiv.org/pdf/1904.05440.pdf)\n\n## Text Generation/NLP\n### Question Generation:\n* [Question Generation: generate multiple choice answers from text](https://github.com/KristiyanVachev/Question-Generation)\n### Other\n* [OpenAI 1.5 billion params GPT-2 release](https://openai.com/blog/gpt-2-1-5b-release/)\n* [AIDungeon](https://github.com/AIDungeon/AIDungeon)\n* [ctrl-gce: CTRL text-generating model on Google Compute Engine with just a few console commands.](https://github.com/minimaxir/ctrl-gce), why google compute engine: The CTRL model is so large (12 GB on disk, 15.5 GB GPU VRAM when loaded, even more system RAM during runtime) that it will currently not fit into a free Colaboratory or Kaggle Notebook. \n* [Writing with the machine](https://www.robinsloan.com/notes/writing-with-the-machine/)\n  [scifi corpus txt dataset](https://www.kaggle.com/jannesklaas/scifi-stories-text-corpus)\n  Could be useful for gpt2 model fine tune\n* [GPT2-Chinese-wuxiao-novel](https://leemeng.tw/gpt2-language-model-generate-chinese-jing-yong-novels.html)\n* [GPT-2 Chinese](https://github.com/Morizeyao/GPT2-Chinese)\n* [Hugging face Transformer]-Pytorch hub \n\n### Text Corpus/dataset\n* [Sci-fi-Script](http://www.scifiscripts.com/)\n* [Detroit-Becoming-Human](https://github.com/detroitbecometext/detroitbecometext.github.io)\n   Could be source to analyze dialogue tree and decision tree structure.\n   \n## Audio and music generation/processing\n### Project:\n* [Talking like your favorite character: Text-To-Speech audio generation ](https://fifteen.ai/app)\nrelated research:\n* [Tacotron2](https://github.com/NVIDIA/tacotron2)\n* [ForwardTacotron: Tacotron2 without attention](https://github.com/as-ideas/ForwardTacotron)\n* Voice clone: [Real-Time-Voice-Cloning](https://github.com/CorentinJ/Real-Time-Voice-Cloning)\n* [Spleeter: sound track seperation](https://github.com/deezer/spleeter)\n  sound seperation is under domain of music information retrival.\n* [Ambient Generative Music by Alex Bainter](https://generative.fm/)\n  Although this project is not generated by algorithm, gives much inspiration in the field of music generation.\n  [Medium](https://medium.com/@metalex9)\n  [code](https://github.com/generative-music/pieces-alex-bainter/blob/master/packages/piece-trees/src/piece.js)\n*[NeuralFunk: Sound design with ML](https://towardsdatascience.com/neuralfunk-combining-deep-learning-with-sound-design-91935759d628)\n\n### Music-Video Synthesis\n* [Deep Music Visualizer using BigGan](https://towardsdatascience.com/the-deep-music-visualizer-using-sound-to-explore-the-latent-space-of-biggan-198cd37dac9a)\n  [code](https://github.com/msieg/deep-music-visualizer)\n\n## Video Synthesis and Generation:\n* [pix2pix-tensorflow: poweed interative rendereed Virtual world](https://www.youtube.com/watch?v=ayPqjPekn7g)\n  [code](https://github.com/affinelayer/pix2pix-tensorflow)\n   update: pretrained model added\n   [more about](https://affinelayer.com/pix2pix/)\n   [colab](https://www.tensorflow.org/tutorials/generative/pix2pix)\n   * Using this technique we can colorize black and white photos, convert google maps to google earth, etc.\n\n* [CRAFT, which generates cartoons based on text descritpionsa](https://arxiv.org/abs/1804.03608)\n  A very creative work involved text to video generation from allen nlp\n  by [researcher](http://tanmaygupta.info/publications/)\n  [Project page](https://prior.allenai.org/projects/craft)\n  [Video](https://www.youtube.com/watch?v=688Vv86n0z8\u0026feature=youtu.be)\n\n## Audio Systhesis\n* [nvidia's taco2 pytorch implementation](https://github.com/NVIDIA/tacotron2)\n* [real-time-voice-clone](https://github.com/CorentinJ/Real-Time-Voice-Cloning)\n\n## ProceduralGenerationForGaming\n### Map generation\n* [AI-Powered Procedural Fantasy Map Generator](https://80.lv/articles/ai-powered-procedural-fantasy-map-generator/), reference: [@linonetwo's blog](https://onetwo.ren/wiki/#:%E5%9C%B0%E5%9B%BE%E7%94%9F%E6%88%90%E7%9A%84%E4%BE%8B%E5%AD%90)\n\n## Animation\n* [deep learning for character animation and control](https://github.com/sebastianstarke/AI4Animation)\n* [DeepMimic: Motion imitation with deep reinforcement learning](https://github.com/xbpeng/DeepMimic)\n\n# Distributed Training:\n* [MPI Reduce and Allreduce](https://mpitutorial.com/tutorials/mpi-reduce-and-allreduce/)\n   Very useful tutorial to illustrate the concept of MPT.\n   CHECK [CODE](https://github.com/wesleykendall/mpitutorial/tree/gh-pages)\n* Also check on [Tensorflow-core](https://www.tensorflow.org/tutorials/distribute/keras)\n* [Distributed Traning strategy](https://www.tensorflow.org/guide/distributed_training)\n\n## Computer Vision application in gaming\n* [E-Sports Talent Scouting Based on Multimodal Twitch Stream Data](https://arxiv.org/pdf/1907.01615.pdf)\n  [data acquisition and modeling code](https://github.com/mug31416/E-sports-on-Twitch/tree/master/data)\n  [Chat Log](https://github.com/bernardopires/twitch-chat-logger)\n  [Twich Stream](https://www.godo.dev/tutorials/python-record-twitch/)\n\n## Paper with code\n* [Paper with code](https://paperswithcode.com/)\n\n## Resources on related course:\n* [UIUC CS598RK: HCI for ML](https://courses.grainger.illinois.edu/cs598rk/fa2019/)\n* [coursera: Sequence Models](https://www.coursera.org/learn/nlp-sequence-models/)\n* [Full Stack Deep Learning](https://fullstackdeeplearning.com/)\n* fast.ai\n* [UIUC: ECE420 Video processing lab](https://courses.grainger.illinois.edu/ece420/fa2019/lab7/lab/)\n* GPU free resource:[Setting Up a Google Cloud Instance GPU for fast.ai for Free](https://medium.com/@jamsawamsa/running-a-google-cloud-gpu-for-fast-ai-for-free-5f89c707bae6)\n\n## Cloud computing set up\n* [AWS](https://sebastianraschka.com/pdf/books/dlb/appendix_cloud-computing.pdf)\n* [Google colab](https://medium.com/@mslavescu/try-live-ssd-object-detection-mask-r-cnn-object-detection-and-instance-segmentation-sfmlearner-df62bdc97d52)\n\n## Full Stack Deep Learning tools in data processing pipeline:\n* [cortex: Deploy machine learning models in production possibly without docker and kubernetes](https://github.com/cortexlabs/cortex)\n   [medium](https://towardsdatascience.com/@calebkaiser)\n\n## Resources for dev tool:\n* Lumen: Video syth software\n* [Dialogue tree based node editor for unity](https://github.com/Seneral/Node_Editor_Framework)\n  and [example](https://github.com/Seneral/Node_Editor_Framework/tree/Examples/Dialogue-System)\n  [demo](https://nodeeditor.seneral.dev/Examples.html)\n\n\n## More resources on pretrained model:\n * [Pytorch hub]https://pytorch.org/hub/research-models\n * Tensorflow hub\n\n## MEET mind-linked people in forum:\n* [reddit-MediaSynthesis](https://www.reddit.com/r/MediaSynthesis/)\n\n## Application:\n### Generating voices of various characters: Ideally podcast and gaming etc:\n* [15.ai](https://15.ai/about)\n\n### FILM MAKING:\n* Black Mirror: Bandersnatch\n  [Show case: dialogue tree](https://www.reddit.com/r/blackmirror/comments/aajk5r/full_bandersnatch_flowchart_all_branches_story/)\n* [LATE-SHIFT](https://lateshift-movie.com/)\n\n## Design\n* [Algorithm Driven Design](https://algorithms.design/)\n* [Algorithm generated Logo](https://app.brandmark.io/)\n\n## Reference:\n* For details about how to represent animation architecture from software perspective\nand in math. Check Chapter11-Animation System of Game Engine Architecture,2nd edition\nby Jason Gregory\n* [Kite's new AI model](https://techcrunch.com/2019/01/28/kite-raises-17m-for-its-ai-driven-code-completion-tool/)\n","projects_url":"https://awesome.ecosyste.ms/api/v1/lists/franceshe%2Fawesome-generative-models/projects"}