An open API service indexing awesome lists of open source software.

Projects in Awesome Lists tagged with image-generation

A curated list of projects in awesome lists tagged with image-generation .

https://github.com/unslothai/unsloth

Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.

agent ai chatgpt deepseek fine-tuning gemma image-generation llama llm llms openai python qwen reinforcement-learning self-hosted stable-diffusion text-to-speech tts ui unsloth

Last synced: 27 Aug 2026

https://github.com/khoj-ai/khoj

Your AI second brain. Self-hostable. Get answers from the web or your docs. Build custom agents, schedule automations, do deep research. Turn any online or local LLM into your personal, autonomous AI (gpt, claude, gemini, llama, qwen, mistral). Get started - free.

agent ai assistant chat chatgpt emacs image-generation llama3 llamacpp llm obsidian obsidian-md offline-llm productivity rag research self-hosted semantic-search stt whatsapp-ai

Last synced: 22 Feb 2026

https://github.com/mudler/localai

:robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video, Images, Voice Cloning, Distributed, P2P inference

ai api audio-generation distributed gemma gpt4all image-generation kubernetes libp2p llama llama3 llm mamba mistral musicgen rerank rwkv stable-diffusion text-generation tts

Last synced: 14 May 2026

https://github.com/go-skynet/LocalAI

:robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video, Images, Voice Cloning, Distributed, P2P inference

ai api audio-generation distributed gemma gpt4all image-generation kubernetes libp2p llama llama3 llm mamba mistral musicgen rerank rwkv stable-diffusion text-generation tts

Last synced: 03 May 2025

https://github.com/invoke-ai/invokeai

Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The solution offers an industry leading WebUI, and serves as the foundation for multiple commercial products.

ai-art artificial-intelligence generative-art image-generation img2img inpainting latent-diffusion linux macos outpainting stable-diffusion txt2img windows

Last synced: 30 Jun 2026

https://github.com/mudler/LocalAI

:robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video, Images, Voice Cloning, Distributed inference

ai api audio-generation distributed gemma gpt4all image-generation kubernetes llama llama3 llm mamba mistral musicgen p2p rerank rwkv stable-diffusion text-generation tts

Last synced: 14 Mar 2025

https://invoke-ai.github.io/InvokeAI/

Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The solution offers an industry leading WebUI, and serves as the foundation for multiple commercial products.

ai-art artificial-intelligence generative-art image-generation img2img inpainting latent-diffusion linux macos outpainting stable-diffusion txt2img windows

Last synced: 08 May 2025

https://github.com/invoke-ai/InvokeAI

InvokeAI is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The solution offers an industry leading WebUI, supports terminal use through a CLI, and serves as the foundation for multiple commercial products.

ai-art artificial-intelligence generative-art image-generation img2img inpainting latent-diffusion linux macos outpainting stable-diffusion txt2img windows

Last synced: 14 Mar 2025

https://github.com/graphiteeditor/graphite

An open source graphics editor for 2025: comprehensive 2D content creation tool suite for graphic design, digital art, and interactive real-time motion graphics — featuring node-based procedural editing

2d-graphics animation art creative-coding design graphic-design graphics graphics-editor image-generation image-manipulation image-processing motion-design motion-graphics node-graph photo-editor procedural procedural-drawing procedural-generation svg-editor vector-graphics

Last synced: 09 Sep 2025

https://github.com/vercel/satori

Enlightened library to convert HTML and CSS to SVG

css image image-generation image-generator jsx opengraph-images satori svg vercel

Last synced: 06 Jul 2026

https://github.com/junyanz/cyclegan

Software that can generate photos from paintings, turn horses into zebras, perform style transfer, and more.

computer-graphics computer-vision cyclegan deep-learning gan gans generative-adversarial-network image-generation image-manipulation pix2pix torch

Last synced: 29 Apr 2025

https://github.com/junyanz/CycleGAN

Software that can generate photos from paintings, turn horses into zebras, perform style transfer, and more.

computer-graphics computer-vision cyclegan deep-learning gan gans generative-adversarial-network image-generation image-manipulation pix2pix torch

Last synced: 13 Mar 2025

https://junyanz.github.io/CycleGAN/

Software that can generate photos from paintings, turn horses into zebras, perform style transfer, and more.

computer-graphics computer-vision cyclegan deep-learning gan gans generative-adversarial-network image-generation image-manipulation pix2pix torch

Last synced: 18 Mar 2025

https://github.com/alexjc/neural-doodle

Turn your two-bit doodles into fine artworks with deep neural networks, generate seamless textures from photos, transfer style from one image to another, perform example-based upscaling, but wait... there's more! (An implementation of Semantic Style Transfer.)

deep-learning deep-neural-networks image-generation image-manipulation image-processing

Last synced: 27 Sep 2025

https://github.com/paddlepaddle/paddlegan

PaddlePaddle GAN library, including lots of interesting applications like First-Order motion transfer, Wav2Lip, picture repair, image editing, photo2cartoon, image style transfer, GPEN, and so on.

animeganv2 basicvsrplusplus cyclegan edvr first-order-motion-model gan gpen image-editing image-generation motion-transfer photo2cartoon pix2pix psgan realsr resolution stylegan2 super-resolution wav2lip

Last synced: 13 May 2025

https://github.com/PaddlePaddle/PaddleGAN

PaddlePaddle GAN library, including lots of interesting applications like First-Order motion transfer, Wav2Lip, picture repair, image editing, photo2cartoon, image style transfer, GPEN, and so on.

animeganv2 basicvsrplusplus cyclegan edvr first-order-motion-model gan gpen image-editing image-generation motion-transfer photo2cartoon pix2pix psgan realsr resolution stylegan2 super-resolution wav2lip

Last synced: 07 Apr 2025

https://github.com/carson-katri/dream-textures

Stable Diffusion built-in to Blender

ai blender blender-addon image-generation stable-diffusion

Last synced: 14 May 2025

https://github.com/foundationvision/var

[NeurIPS 2024 Best Paper][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction". An *ultra-simple, user-friendly yet state-of-the-art* codebase for autoregressive image generation!

auto-regressive-model autoregressive-models diffusion-models generative-ai generative-model gpt gpt-2 image-generation large-language-models neurips transformers vision-transformer

Last synced: 10 Apr 2025

https://github.com/open-mmlab/mmagic

OpenMMLab Multimodal Advanced, Generative, and Intelligent Creation Toolbox. Unlock the magic 🪄: Generative-AI (AIGC), easy-to-use APIs, awsome model zoo, diffusion models, for text-to-image generation, image/video restoration/enhancement, etc.

aigc computer-vision deep-learning diffusion diffusion-models generative-adversarial-network generative-ai image-editing image-generation image-processing image-synthesis inpainting matting pytorch super-resolution text2image video-frame-interpolation video-interpolation video-super-resolution

Last synced: 13 Dec 2025

https://github.com/yuanxiaosc/DeepImage-an-Image-to-Image-technology

DeepNude's algorithm and general image generation theory and practice research, including pix2pix, CycleGAN, UGATIT, DCGAN, SinGAN, ALAE, mGANprior, StarGAN-v2 and VAE models (TensorFlow2 implementation). DeepNude的算法以及通用生成对抗网络(GAN,Generative Adversarial Network)图像生成的理论与实践研究。

cycle-gan dcgan deepface deepfakes deepnude image-generation image-to-image nerual-style pix2pix sin-gan style-gan tensorflow2 vae zao

Last synced: 26 Mar 2025

https://github.com/OpenGVLab/DragGAN

Unofficial Implementation of DragGAN - "Drag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifold" (DragGAN 全功能实现,在线Demo,本地部署试用,代码、模型已全部开源,支持Windows, macOS, Linux)

draggan gradio-interface image-editing image-generation interngpt

Last synced: 02 Apr 2025

https://github.com/opengvlab/draggan

Unofficial Implementation of DragGAN - "Drag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifold" (DragGAN 全功能实现,在线Demo,本地部署试用,代码、模型已全部开源,支持Windows, macOS, Linux)

draggan gradio-interface image-editing image-generation interngpt

Last synced: 14 May 2025

https://github.com/stability-ai/stableswarmui

StableSwarmUI, A Modular Stable Diffusion Web-User-Interface, with an emphasis on making powertools easily accessible, high performance, and extensibility.

ai image-generation stable-diffusion stablediffusion ui

Last synced: 13 May 2025

https://github.com/Stability-AI/StableSwarmUI

StableSwarmUI, A Modular Stable Diffusion Web-User-Interface, with an emphasis on making powertools easily accessible, high performance, and extensibility.

ai image-generation stable-diffusion stablediffusion ui

Last synced: 28 Mar 2025

https://github.com/knplabs/snappy

PHP library allowing thumbnail, snapshot or PDF generation from a url or a html page. Wrapper for wkhtmltopdf/wkhtmltoimage

hacktoberfest html-to-image html-to-pdf image-generation pdf-generation php

Last synced: 12 May 2025

https://github.com/KnpLabs/snappy

PHP library allowing thumbnail, snapshot or PDF generation from a url or a html page. Wrapper for wkhtmltopdf/wkhtmltoimage

hacktoberfest html-to-image html-to-pdf image-generation pdf-generation php

Last synced: 14 Mar 2025

https://github.com/FoundationVision/VAR

[NeurIPS 2024 Oral][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction". An *ultra-simple, user-friendly yet state-of-the-art* codebase for autoregressive image generation!

auto-regressive-model autoregressive-models diffusion-models generative-ai generative-model gpt gpt-2 image-generation large-language-models neurips transformers vision-transformer

Last synced: 03 Apr 2025

https://github.com/calesthio/OpenMontage

World's first open-source, agentic video production system. 12 pipelines, 52 tools, 500+ agent skills. Turn your AI coding assistant into a full video production studio.

agent agentic-ai ai claude copilot cursor elevenlabs ffmpeg flux image-generation open-source openai python remotion stable-diffusion text-to-speech text-to-video video-generation video-production

Last synced: 22 Jun 2026

https://github.com/janspiry/image-super-resolution-via-iterative-refinement

Unofficial implementation of Image Super-Resolution via Iterative Refinement by Pytorch

ddpm diffusion-probabilistic image-generation pytorch super-resolution

Last synced: 14 May 2025

https://github.com/s1dashu/ip-as-logo-skill

A compact Agent Skill for highly simplified, rounded, subtly neo-skeuomorphic IP mascot logos.

codex codex-skill image-generation logo-design mascot-design

Last synced: 24 Aug 2026

https://github.com/ali-vilab/AnyDoor

Official implementations for paper: Anydoor: zero-shot object-level image customization

image-composition image-customization image-editing image-generation

Last synced: 28 Mar 2025

https://github.com/Janspiry/Image-Super-Resolution-via-Iterative-Refinement

Unofficial implementation of Image Super-Resolution via Iterative Refinement by Pytorch

ddpm diffusion-probabilistic image-generation pytorch super-resolution

Last synced: 25 Mar 2025

https://github.com/AIDC-AI/Pixelle-Video

🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine

aigc comfyui image-generation tts video-generation

Last synced: 03 May 2026

https://github.com/joepenna/dreambooth-stable-diffusion

Implementation of Dreambooth (https://arxiv.org/abs/2208.12242) by way of Textual Inversion (https://arxiv.org/abs/2208.01618) for Stable Diffusion (https://arxiv.org/abs/2112.10752). Tweaks focused on training faces, objects, and styles.

ai artificial-intelligence image-generation img2img latent-diffusion machine-learning model-training stable-diffusion txt2img

Last synced: 14 May 2025

https://github.com/JoePenna/Dreambooth-Stable-Diffusion

Implementation of Dreambooth (https://arxiv.org/abs/2208.12242) by way of Textual Inversion (https://arxiv.org/abs/2208.01618) for Stable Diffusion (https://arxiv.org/abs/2112.10752). Tweaks focused on training faces, objects, and styles.

ai artificial-intelligence image-generation img2img latent-diffusion machine-learning model-training stable-diffusion txt2img

Last synced: 16 Apr 2025

https://github.com/CookSleep/gpt_image_playground

基于 OpenAI gpt-image-2 API 的图片生成与编辑工具

gpt-image image-editing image-generation openai react tailwindcss typescript vite

Last synced: 14 Aug 2026

https://github.com/elegantapp/pwa-asset-generator

Automates PWA asset generation and image declaration. Automatically generates icon and splash screen images, favicons and mstile images. Updates manifest.json and index.html files with the generated images according to Web App Manifest specs and Apple Human Interface guidelines.

android favicon favicon-generator html-tags icon icon-sizes image-generation ios launch-image manifest manifest-specs mstile progressive-web-app progressive-web-apps puppeteer pwa pwa-assets splash-screen splash-screens

Last synced: 14 Feb 2026

https://github.com/ai-forever/kandinsky-2

Kandinsky 2 — multilingual text2image latent diffusion model

diffusion image-generation image2image inpainting ipython-notebook kandinsky outpainting text-to-image text2image

Last synced: 15 May 2025

https://github.com/ai-forever/Kandinsky-2

Kandinsky 2 — multilingual text2image latent diffusion model

diffusion image-generation image2image inpainting ipython-notebook kandinsky outpainting text-to-image text2image

Last synced: 08 Apr 2025

https://github.com/mcmonkeyprojects/swarmui

SwarmUI (formerly StableSwarmUI), A Modular Stable Diffusion Web-User-Interface, with an emphasis on making powertools easily accessible, high performance, and extensibility.

ai comfyui csharp image-generation javascript machine-learning ml python stable-diffusion

Last synced: 14 May 2025

https://github.com/taesungp/contrastive-unpaired-translation

Contrastive unpaired image-to-image translation, faster and lighter training than cyclegan (ECCV 2020, in PyTorch)

computer-graphics computer-vision computervision cyclegan deeplearning gans generative-adversarial-network image-generation image-manipulation pytorch

Last synced: 15 May 2025

https://github.com/Anil-matcha/awesome-generative-ai-apps

50+ open-source generative AI apps you can clone, deploy, and monetize — image generators, video tools, virtual try-ons, AI SaaS templates, and platform integrations. One-click Vercel deploy on every template.

ai-apps ai-image-generator ai-saas ai-starter ai-tools ai-video-generator awesome awesome-list generative-ai generative-ai-apps image-generation nextjs nextjs-boilerplate open-source open-source-ai saas-boilerplate saas-template text-to-image text-to-video video-generation

Last synced: 26 Jun 2026

https://github.com/crmne/ruby_llm

Stop juggling AI SDKs! RubyLLM offers one delightful Ruby interface for OpenAI, Anthropic, Gemini, Bedrock, OpenRouter, DeepSeek, Ollama & compatible APIs. Chat, Vision, Audio, PDF, Images, Embeddings, Tools, Streaming & Rails integration.

ai anthropic chatgpt claude dall-e deepseek embeddings gemini image-generation llm openai rails ruby

Last synced: 02 Apr 2026

https://github.com/kane50613/takumi

Render JSX, HTML, and CSS to images without a headless browser. OG cards, animated GIFs, and video frames from Node.js, edge runtimes, browsers, or Rust. Drop-in next/og replacement.

cloudflare-workers css gif html-to-image image-generation jsx nodejs og-image opengraph react rust satori tailwindcss wasm

Last synced: 12 Jun 2026

https://github.com/foundationvision/llamagen

Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation

auto-regressive-model diffusion diffusion-models image-generation llama llm text2image

Last synced: 15 May 2025

https://github.com/FoundationVision/LlamaGen

Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation

auto-regressive-model diffusion diffusion-models image-generation llama llm text2image

Last synced: 07 May 2025

https://github.com/pydn/comfyui-to-python-extension

A powerful tool that translates ComfyUI workflows into executable Python code.

ai-art comfyui generative-art image-generation pytorch stable-diffusion

Last synced: 02 Apr 2026

https://github.com/varunshenoy/opendream

An extensible, easy-to-use, and portable diffusion web UI 👨‍🎨

ai automatic-1111 diffusion image-generation stable-diffusion

Last synced: 08 Apr 2025

https://github.com/minimax-ai/minimax-mcp

Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech, image generation and video generation APIs.

image-generation image-to-video mcp mcp-server mcp-tools text-to-image text-to-speech text-to-video video-generation voice-cloning

Last synced: 28 Aug 2026

https://github.com/Anil-matcha/Awesome-GPT-Image-2-API-Prompts

Curated GPT-Image-2 prompts for the OpenAI API — portraits, posters, UI mockups, game screenshots, character sheets, and more. Ready-to-use prompts for gpt-image-2.

ai-art ai-generated-art ai-image api awesome-list chatgpt dalle generative-ai gpt-image-2 gpt-image-2-prompts image-generation image-prompts midjourney-prompts openai openai-api prompt-collection prompt-engineering prompts stable-diffusion-prompts text-to-image

Last synced: 04 May 2026

https://github.com/pydn/ComfyUI-to-Python-Extension

A powerful tool that translates ComfyUI workflows into executable Python code.

ai-art comfyui generative-art image-generation pytorch stable-diffusion

Last synced: 19 Aug 2025

https://github.com/mit-han-lab/data-efficient-gans

[NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training

data-efficient gans generative-adversarial-network image-generation neurips-2020 pytorch tensorflow

Last synced: 15 May 2025

https://github.com/poloclub/diffusiondb

A large-scale text-to-image prompt gallery dataset based on Stable Diffusion

ai-art computer-vision image-generation prompt-engineering stable-diffusion

Last synced: 16 May 2025

https://poloclub.github.io/diffusiondb/

A large-scale text-to-image prompt gallery dataset based on Stable Diffusion

ai-art computer-vision image-generation prompt-engineering stable-diffusion

Last synced: 07 May 2025

https://github.com/KnpLabs/KnpSnappyBundle

Easily create PDF and images in Symfony by converting html using webkit

hacktoberfest html-to-image html-to-pdf image-generation pdf-generation php symfony symfony-bundle

Last synced: 31 Mar 2025

https://github.com/knplabs/knpsnappybundle

Easily create PDF and images in Symfony by converting html using webkit

hacktoberfest html-to-image html-to-pdf image-generation pdf-generation php symfony symfony-bundle

Last synced: 14 May 2025

https://github.com/bytedance/Lance

A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.

image-editing image-generation image-understanding unified-multimodal-models video-generation video-understanding

Last synced: 26 Jun 2026

https://github.com/receyuki/stable-diffusion-prompt-reader

A simple standalone viewer for reading prompts from Stable Diffusion generated image outside the webui.

ai image-generation macos prompt-engineering prompt-toolkit python stable-diffusion stable-diffusion-webui tkinter windows

Last synced: 15 May 2025

https://github.com/0xacx/chatGPT-shell-cli

Simple shell script to use OpenAI's ChatGPT and DALL-E from the terminal. No Python or JS required.

bash chatbot chatgpt chatgpt-api chatgpt-api-wrapper cli dall-e dalle dalle2 image-generation shell shell-script terminal zsh

Last synced: 14 Mar 2025

https://github.com/0xacx/chatgpt-shell-cli

Simple shell script to use OpenAI's ChatGPT and DALL-E from the terminal. No Python or JS required.

bash chatbot chatgpt chatgpt-api chatgpt-api-wrapper cli dall-e dalle dalle2 image-generation shell shell-script terminal zsh

Last synced: 08 Apr 2025

https://github.com/ermongroup/sdedit

PyTorch implementation for SDEdit: Image Synthesis and Editing with Stochastic Differential Equations

controllable-generation image-editing image-generation image-manipulation pytorch score-matching

Last synced: 16 May 2025

https://github.com/omerbt/multidiffusion

Official Pytorch Implementation for "MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation" presenting "MultiDiffusion" (ICML 2023)

diffusion-models generative-model icml image-generation multidiffusion stable-diffusion text-to-image

Last synced: 16 May 2025

https://github.com/aws-samples/generative-ai-use-cases

Application implementation with business use cases for safely utilizing generative AI in business operations

aws bedrock chatbot claude claude3 command-r deepseek-r1 generative-ai image-generation lambda llama3 llm mistral nova rag react sagemaker typescript

Last synced: 17 Nov 2025

https://github.com/ermongroup/SDEdit

PyTorch implementation for SDEdit: Image Synthesis and Editing with Stochastic Differential Equations

controllable-generation image-editing image-generation image-manipulation pytorch score-matching

Last synced: 28 Mar 2025

https://github.com/thudm/cogview4

CogView4, CogView3-Plus and CogView3(ECCV 2024)

eccv2024 high-resolution image-generation text-to-image

Last synced: 15 May 2025

https://github.com/omerbt/MultiDiffusion

Official Pytorch Implementation for "MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation" presenting "MultiDiffusion" (ICML 2023)

diffusion-models generative-model icml image-generation multidiffusion stable-diffusion text-to-image

Last synced: 28 Mar 2025

https://github.com/THUDM/CogView4

CogView4, CogView3-Plus and CogView3(ECCV 2024)

eccv2024 high-resolution image-generation text-to-image

Last synced: 01 Apr 2025

https://github.com/UCSC-VLAA/story-iter

[ICLR 2026] A Training-free Iterative Framework for Long Story Visualization

diffusion-models generative-art generative-model image-generation storytelling visual-storytelling

Last synced: 08 Feb 2026

https://github.com/cure-lab/magicdrive

[ICLR24] Official implementation of the paper “MagicDrive: Street View Generation with Diverse 3D Geometry Control”

autonomous-vehicles deep-learning diffusion-models image-generation pytorch video-generation

Last synced: 15 May 2025

https://github.com/bryandlee/malnyun_faces

침착한 생성모델 학습기

image-generation image-to-image-translation

Last synced: 02 Feb 2026

https://github.com/yunjey/domain-transfer-network

TensorFlow Implementation of Unsupervised Cross-Domain Image Generation

domain-transfer image-generation tensorflow unsupervised-learning

Last synced: 24 Mar 2025

https://github.com/aws-samples/generative-ai-use-cases-jp

すぐに業務活用できるビジネスユースケース集付きの安全な生成AIアプリ実装

aws bedrock chatbot claude claude3 command-r deepseek-r1 generative-ai image-generation lambda llama3 llm mistral nova rag react sagemaker typescript

Last synced: 30 Mar 2025