Projects in Awesome Lists tagged with video-generation
A curated list of projects in awesome lists tagged with video-generation .
https://github.com/KlingAIResearch/LivePortrait
Bring portraits to life!
face-animation image-animation video-editing video-generation
Last synced: 20 Feb 2026
https://github.com/kwaivgi/liveportrait
Bring portraits to life!
face-animation image-animation video-editing video-generation
Last synced: 14 May 2025
https://github.com/KwaiVGI/LivePortrait
Bring portraits to life!
face-animation image-animation video-editing video-generation
Last synced: 26 Mar 2025
https://github.com/waooAI/waoowaoo
首家工业级全流程 AI 影视生产平台。Industry-first professional AI Agent platform for controllable film & video production. From shorts to live-action with Hollywood-standard workflows.
ai-agent ai-agents automation film-production generative-ai short-drama storyboard video-generation
Last synced: 14 May 2026
https://github.com/zai-org/CogVideo
text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
cogvideox image-to-video llm sora text-to-video video-generation
Last synced: 30 Jul 2025
https://github.com/thudm/cogvideo
text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
cogvideox image-to-video llm sora text-to-video video-generation
Last synced: 16 May 2025
https://github.com/THUDM/CogVideo
text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
cogvideox image-to-video llm sora text-to-video video-generation
Last synced: 28 Mar 2025
https://github.com/dramaclaw/dramaclaw
A general-purpose AIGC video engine: script to finished film in one pipeline — dramas, ads, product videos, otome games, and more. | 通用 AIGC 视频引擎 —— 从剧本到成片一条流水线,漫剧、广告、电商、乙游皆可
ai-agent ai-filmmaking ai-video aigc aigc-pipeline content-creation dramaclaw fastapi generative-ai image-to-video python react self-hosted storyboard text-to-video tts video-engine video-generation voice-synthesis
Last synced: 12 Sep 2026
https://github.com/ailab-cvc/videocrafter
VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
image-to-video text-to-video video-generation
Last synced: 14 May 2025
https://ailab-cvc.github.io/videocrafter/
VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
image-to-video text-to-video video-generation
Last synced: 28 Mar 2025
https://github.com/AILab-CVC/VideoCrafter
VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
image-to-video text-to-video video-generation
Last synced: 28 Mar 2025
https://github.com/fudan-generative-vision/champ
Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance
human-animation image-animatioln video-generation
Last synced: 14 May 2025
https://github.com/picsart-ai-research/text2video-zero
[ICCV 2023 Oral] Text-to-Image Diffusion Models are Zero-Shot Video Generators
video-editing video-generation
Last synced: 13 Apr 2025
https://github.com/Picsart-AI-Research/Text2Video-Zero
[ICCV 2023 Oral] Text-to-Image Diffusion Models are Zero-Shot Video Generators
video-editing video-generation
Last synced: 28 Mar 2025
https://github.com/calesthio/OpenMontage
World's first open-source, agentic video production system. 12 pipelines, 52 tools, 500+ agent skills. Turn your AI coding assistant into a full video production studio.
agent agentic-ai ai claude copilot cursor elevenlabs ffmpeg flux image-generation open-source openai python remotion stable-diffusion text-to-speech text-to-video video-generation video-production
Last synced: 22 Jun 2026
https://github.com/AIDC-AI/Pixelle-Video
🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine
aigc comfyui image-generation tts video-generation
Last synced: 03 May 2026
https://github.com/nexu-io/html-video
Programmatic video for coding agents — HTML to video on your laptop. Turn HTML, CSS & data into real MP4s with pluggable render engines, 21 templates, AI soundtrack. Apache-2.0, no per-render fees. An official project by the Open Design team.
ai-agent apache-2 coding-agent css ffmpeg html html-to-video hyperframes mp4 open-design open-source programmatic-video video video-as-code video-generation
Last synced: 20 Jul 2026
https://github.com/OpenGVLab/InternGPT
InternGPT (iGPT) is an open source demo platform where you can easily showcase your AI models. Now it supports DragGAN, ChatGPT, ImageBind, multimodal chat like GPT-4, SAM, interactive image editing, etc. Try it at igpt.opengvlab.com (支持DragGAN、ChatGPT、ImageBind、SAM的在线Demo系统)
chatgpt click draggan foundation-model gpt gpt-4 gradio husky image-captioning imagebind internimage langchain llama llm multimodal sam segment-anything vicuna video-generation vqa
Last synced: 27 Mar 2025
https://github.com/opengvlab/interngpt
InternGPT (iGPT) is an open source demo platform where you can easily showcase your AI models. Now it supports DragGAN, ChatGPT, ImageBind, multimodal chat like GPT-4, SAM, interactive image editing, etc. Try it at igpt.opengvlab.com (支持DragGAN、ChatGPT、ImageBind、SAM的在线Demo系统)
chatgpt click draggan foundation-model gpt gpt-4 gradio husky image-captioning imagebind internimage langchain llama llm multimodal sam segment-anything vicuna video-generation vqa
Last synced: 14 May 2025
https://github.com/SandAI-org/MAGI-1
MAGI-1: Autoregressive Video Generation at Scale
autoregressive diffusion-models video-generation
Last synced: 13 Jun 2025
https://github.com/ArcReel/ArcReel
AI Agent 驱动的开源视频生成工作台 — 小说→角色/场景/道具设计→剧本→分镜图→视频,跨镜头角色与场景一致 | Open-source AI video workspace powered by AI Agents, Nano Banana 2 & Veo 3.1 / Grok / Seedance / OpenAI
ai-agent ai-video-generator claude-agent-sdk docker gemini grok image-to-video nano-banana-2 openai openclaw seedance seedream storyboard veo vertex-ai video-generation
Last synced: 11 Jul 2026
https://github.com/Vincentwei1021/video-shotcraft
AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template
agent-skills ai-agents ai-video claude-code claude-code-skills claude-skills codex motion-design motion-graphics product-video promo-video remotion video-generation video-production
Last synced: 04 Aug 2026
https://github.com/doubiiu/dynamicrafter
[ECCV 2024, Oral] DynamiCrafter: Animating Open-domain Images with Video Diffusion Priors
image-animation image-to-video video-generation
Last synced: 15 May 2025
https://github.com/Doubiiu/DynamiCrafter
[ECCV 2024, Oral] DynamiCrafter: Animating Open-domain Images with Video Diffusion Priors
image-animation image-to-video video-generation
Last synced: 14 Apr 2025
https://github.com/mutonby/openshorts
Free & open source AI video platform — Clip Generator, AI Shorts (UGC with AI actors) & YouTube Studio. Self-hosted, no watermarks.
clip-generator open-source shorts tiktok ugc-platform video-editing video-generation
Last synced: 25 Jul 2026
https://github.com/tmelyralab/musev
MuseV: Infinite-length and High Fidelity Virtual Human Video Generation with Visual Conditioned Parallel Denoising
diffusion human-video-generation image2video infinite-length musev video-generation
Last synced: 26 Sep 2025
https://github.com/TMElyralab/MuseV
MuseV: Infinite-length and High Fidelity Virtual Human Video Generation with Visual Conditioned Parallel Denoising
diffusion human-video-generation image2video infinite-length musev video-generation
Last synced: 11 Apr 2025
https://github.com/tencent/MimicMotion
High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
diffusion-models video-generation
Last synced: 11 May 2025
https://github.com/tencent/mimicmotion
High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
diffusion-models video-generation
Last synced: 15 May 2025
https://github.com/HKUDS/ViMax
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
Last synced: 31 Mar 2026
https://github.com/Anil-matcha/awesome-generative-ai-apps
50+ open-source generative AI apps you can clone, deploy, and monetize — image generators, video tools, virtual try-ons, AI SaaS templates, and platform integrations. One-click Vercel deploy on every template.
ai-apps ai-image-generator ai-saas ai-starter ai-tools ai-video-generator awesome awesome-list generative-ai generative-ai-apps image-generation nextjs nextjs-boilerplate open-source open-source-ai saas-boilerplate saas-template text-to-image text-to-video video-generation
Last synced: 26 Jun 2026
https://github.com/samuraigpt/ai-youtube-shorts-generator
A python tool that uses GPT-4, FFmpeg, and OpenCV to automatically analyze videos, extract the most interesting sections, and crop them for an improved viewing experience.
ai-video-generator artificial-intelligence image-to-video image-to-video-generation shorts shorts-maker sora-video sora-video-ai stable-diffusion text-to-image text-to-video text-to-video-generation video-diffusion video-editing video-generation video-generator youtube-shorts
Last synced: 14 May 2025
https://github.com/ali-vilab/dreamtalk
Official implementations for paper: DreamTalk: When Expressive Talking Head Generation Meets Diffusion Probabilistic Models
audio-visual-learning face-animation talking-head video-generation
Last synced: 15 May 2025
https://github.com/microsoft/researchstudio
ResearchStudio: Our AI co-author, from research problem to final publication.
blog-generation idea-generation poster-generation presentation researchstudio-idea researchstudio-reel video-generation
Last synced: 23 Jul 2026
https://github.com/minimax-ai/minimax-mcp
Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech, image generation and video generation APIs.
image-generation image-to-video mcp mcp-server mcp-tools text-to-image text-to-speech text-to-video video-generation voice-cloning
Last synced: 28 Aug 2026
https://github.com/thu-ml/sageattention
Quantized Attention achieves speedup of 2-3x and 3-5x compared to FlashAttention and xformers, without lossing end-to-end metrics across language, image, and video models.
attention cuda efficient-attention inference-acceleration llm llm-infra mlsys quantization triton video-generate video-generation vit
Last synced: 14 May 2025
https://github.com/skyworkai/matrix-game
Matrix-Game 2.0: An Open-Source, Real-Time, and Streaming Interactive World Model
genie interactive-video long-sequence long-video real-time video-generation world-model
Last synced: 23 Sep 2025
https://pku-yuangroup.github.io/MagicTime/
[TPAMI 2025🔥] MagicTime: Time-lapse Video Generation Models as Metamorphic Simulators
diffusion-models long-video-generation metamorphic-video-generation open-sora-plan text-to-video time-lapse time-lapse-dataset video-generation
Last synced: 23 Aug 2026
https://github.com/mayuelala/followyourpose
[AAAI 2024] Follow-Your-Pose: This repo is the official implementation of "Follow-Your-Pose : Pose-Guided Text-to-Video Generation using Pose-Free Videos"
aaai-2024 aigc follow-your-pose laion-pose-dataset video-generation
Last synced: 16 May 2025
https://github.com/mayuelala/FollowYourPose
[AAAI 2024] Follow-Your-Pose: This repo is the official implementation of "Follow-Your-Pose : Pose-Guided Text-to-Video Generation using Pose-Free Videos"
aaai-2024 aigc follow-your-pose laion-pose-dataset video-generation
Last synced: 28 Mar 2025
https://github.com/lucidrains/video-diffusion-pytorch
Implementation of Video Diffusion Models, Jonathan Ho's new paper extending DDPMs to Video Generation - in Pytorch
artificial-intelligence ddpm deep-learning text-to-video video-generation
Last synced: 14 Apr 2025
https://github.com/mini-sora/minisora
MiniSora: A community aims to explore the implementation path and future development direction of Sora.
diffusion sora video-generation
Last synced: 14 May 2025
https://github.com/snap-research/articulated-animation
Code for Motion Representations for Articulated Animation paper
deep-learning first-order-motion-model image-animation video-generation
Last synced: 15 May 2025
https://github.com/bytedance/Lance
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
image-editing image-generation image-understanding unified-multimodal-models video-generation video-understanding
Last synced: 26 Jun 2026
https://github.com/wladradchenko/wunjo.wladradchenko.ru
Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.
controlnet deepfake diffusion-models face-animation face-swap free img2video lip-sync photo-editing public-api remove-background remover restyle segment-anything txt2video video-editing video-generation voice-clone wunjo
Last synced: 30 Apr 2026
https://github.com/ali-vilab/UniAnimate
Code for SCIS-2025 Paper "UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation".
dance-generation human-image-animation video-generation
Last synced: 11 May 2025
https://github.com/google-research/magvit
Official JAX implementation of MAGVIT: Masked Generative Video Transformer
generative-model transformers video-generation
Last synced: 27 Mar 2025
https://github.com/cuixing158/Awesome-CV-MasterHub
:fire: :fire: :fire: A paper list of some recent Computer Vision(CV) works
awesome image-captioning image-classification image-dehazing image-denoising image-enhancement image-fusion image-generation image-segmentation keypoint-detection low-level-vision object-detection panoptic-segmentation paper-code paper-list papers-with-code pose-estimation video-generation video-understanding vision-transformer
Last synced: 12 Sep 2026
https://github.com/Vchitect/SEINE
[ICLR 2024] SEINE: Short-to-Long Video Diffusion Model for Generative Transition and Prediction
diffusion-model image-to-video stable-diffusion video-generation
Last synced: 22 Jul 2025
https://github.com/cure-lab/magicdrive
[ICLR24] Official implementation of the paper “MagicDrive: Street View Generation with Diverse 3D Geometry Control”
autonomous-vehicles deep-learning diffusion-models image-generation pytorch video-generation
Last synced: 15 May 2025
https://github.com/showlab/motiondirector
[ECCV 2024 Oral] MotionDirector: Motion Customization of Text-to-Video Diffusion Models.
diffusion-models motion-customization text-to-motion text-to-video text-to-video-generation video-generation
Last synced: 12 Apr 2025
https://github.com/vchitect/seine
[ICLR 2024] SEINE: Short-to-Long Video Diffusion Model for Generative Transition and Prediction
diffusion-model image-to-video stable-diffusion video-generation
Last synced: 15 May 2025
https://github.com/showlab/MotionDirector
[ECCV 2024 Oral] MotionDirector: Motion Customization of Text-to-Video Diffusion Models.
diffusion-models motion-customization text-to-motion text-to-video text-to-video-generation video-generation
Last synced: 28 Mar 2025
https://showlab.github.io/MotionDirector/
[ECCV 2024 Oral] MotionDirector: Motion Customization of Text-to-Video Diffusion Models.
diffusion-models motion-customization text-to-motion text-to-video text-to-video-generation video-generation
Last synced: 28 Mar 2025
https://github.com/mayuelala/followyourclick
[AAAI 2025] Follow-Your-Click: This repo is the official implementation of "Follow-Your-Click: Open-domain Regional Image Animation via Short Prompts"
image-animation image-to-video-generation video-generation
Last synced: 31 Oct 2025
https://github.com/mayuelala/FollowYourClick
[AAAI 2025] Follow-Your-Click: This repo is the official implementation of "Follow-Your-Click: Open-domain Regional Image Animation via Short Prompts"
image-animation image-to-video-generation video-generation
Last synced: 07 May 2025
https://github.com/vchitect/vbench
[CVPR2024 Highlight] VBench - We Evaluate Video Generation
aigc benchmark dataset evaluation-kit gen-ai stable-diffusion text-to-video video-generation
Last synced: 12 Apr 2025
https://github.com/Stonewuu/ai-fusion-video
【融光】 - 基于 Agent 的全流程AI短剧/漫剧/视频创作平台
agents automation java short-drama video-generation
Last synced: 17 Jun 2026
https://liewfeng.github.io/TeaCache/
Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model
cogvideox diffusion-models hunyuan-video inference-acceleration latte open-sora open-sora-plan video-generation
Last synced: 01 Aug 2025
https://github.com/ali-vilab/videocomposer
Official repo for VideoComposer: Compositional Video Synthesis with Motion Controllability
diffusion-models video-generation video-synthesiswith
Last synced: 29 Dec 2025
https://github.com/HA6Bots/TikTok-Compilation-Video-Generator
A system of bots that collects clips automatically via custom made filters, lets you easily browse these clips, and puts them together into a compilation video ready to be uploaded straight to any social media platform. Full VPS support is provided, along with an accounts system so multiple users can use the bot at once. This bot is split up into three separate programs. The server. The client. The video generator. These programs perform different functions that when combined creates a very powerful system for auto generating compilation videos.
automatic tiktok tiktok-api tiktok-downloader tiktok-scraper video-generation video-processing videoeditor
Last synced: 13 Apr 2025
https://github.com/nvidia-cosmos/cosmos-predict2.5
Cosmos-Predict2.5, the latest version of the Cosmos World Foundation Models (WFMs) family, specialized for simulating and predicting the future state of the world in the form of video.
foundational-models video-generation world-models
Last synced: 24 Feb 2026
https://github.com/ChaofanTao/Autoregressive-Models-in-Vision-Survey?tab=readme-ov-file
[TMLR 2025🔥] A survey for the autoregressive models in vision.
acceleration autoregressive computer-vision deep-learning diffusion embodied-ai image-generation medical-ai motion-prediction multimodal point-cloud survey text-to-image video-generation
Last synced: 23 Jul 2026
https://github.com/ha6bots/tiktok-compilation-video-generator
A system of bots that collects clips automatically via custom made filters, lets you easily browse these clips, and puts them together into a compilation video ready to be uploaded straight to any social media platform. Full VPS support is provided, along with an accounts system so multiple users can use the bot at once. This bot is split up into three separate programs. The server. The client. The video generator. These programs perform different functions that when combined creates a very powerful system for auto generating compilation videos.
automatic tiktok tiktok-api tiktok-downloader tiktok-scraper video-generation video-processing videoeditor
Last synced: 16 May 2025
https://github.com/AlonzoLeeeooo/awesome-video-generation
A collection of awesome video generation studies.
diffusion-models generative-adversarial-networks paper-list text-to-video-generation video-generation
Last synced: 12 Sep 2026
https://github.com/xuanyustudio/LocalMiniDrama
🎬 seedance2接入 开源本地 AI 短剧 & 漫剧生成工具 —— 从故事到成片一站式完成,数据不出本机,短剧工作流管理平台,高灵活度,AI真人剧,AI漫剧本地搞定。 Open-source local AI short drama maker: story → storyboard → video, fully offline, your data stays yours. 纳米流水线
ai ai-agent ai-mini-drama ai-video electron javascript mini-drama nodejs offline offline-short-drama open-source short-video short-video-maker storyboard video-generation vue vue3
Last synced: 11 Jul 2026
https://github.com/alibaba/animate-anything
Fine-Grained Open Domain Image Animation with Motion Guidance
animation video-diffusion-model video-generation
Last synced: 14 Oct 2025
https://github.com/ChaofanTao/Autoregressive-Models-in-Vision-Survey
[TMLR 2025🔥] A survey for the autoregressive models in vision.
acceleration autoregressive computer-vision deep-learning diffusion embodied-ai image-generation medical-ai motion-prediction multimodal point-cloud survey text-to-image video-generation
Last synced: 11 Jun 2026
https://github.com/antgroup/echomimic_v3
[AAAI 2026] EchoMimicV3: 1.3B Parameters are All You Need for Unified Multi-Modal and Multi-Task Human Animation
audio-driven-body-animation audio-driven-portrait-animations human-animation video-generation
Last synced: 08 Feb 2026
https://github.com/opendrivelab/vista
[NeurIPS 2024] A Generalizable World Model for Autonomous Driving
autonomous-driving video-generation world-model
Last synced: 15 May 2025
https://github.com/opendrivelab/driveagi
[CVPR 2024 Highlight] GenAD: Generalized Predictive Model for Autonomous Driving & Foundation Models in Autonomous System
autonomous-driving embodied-ai foundation-model general-artificial-intelligence large-dataset policy-learning video-dataset video-generation world-models
Last synced: 15 May 2025
https://github.com/pku-yuangroup/consisid
[CVPR 2025 Highlight🔥] Identity-Preserving Text-to-Video Generation by Frequency Decomposition
diffusion diffusion-models identity-preserving text-to-video video-generation video-generation-dataset video-generator videogeneration
Last synced: 06 Jul 2025
https://boese0601.github.io/magicdance/
[ICML 2024] MagicPose(also known as MagicDance): Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion
behavior-generation cartoon-animation diffusion-models generative-ai generative-model image-editing video-editing video-generation
Last synced: 28 Mar 2025
https://github.com/Boese0601/MagicDance
[ICML 2024] MagicPose(also known as MagicDance): Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion
behavior-generation cartoon-animation diffusion-models generative-ai generative-model image-editing video-editing video-generation
Last synced: 28 Mar 2025
https://github.com/gracezhao1997/Awesome-Video-World-Models-with-AR-Diffusion
A Curated List of Awesome Video World Models with AR Diffusion: Covering Algorithms, Applications, and Infrastructure, Aimed at Serving as a Comprehensive Resource for Researchers, Practitioners, and Enthusiasts.
ar-diffusion autoregressive awesome-list computer-vision diffusion-models generative-ai video-generation world-models
Last synced: 12 Sep 2026
https://github.com/microsoft/ResearchStudio
ResearchStudio: Our AI co-author, from research problem to final publication.
blog-generation idea-generation poster-generation presentation researchstudio-idea researchstudio-reel video-generation
Last synced: 23 Jul 2026
https://github.com/thu-ml/SageAttention
Quantized Attention that achieves speedups of 2.1-3.1x and 2.7-5.1x compared to FlashAttention2 and xformers, respectively, without lossing end-to-end metrics across various models.
attention cuda inference-acceleration llm quantization triton video-generation
Last synced: 15 Aug 2025
https://github.com/thu-ml/riflex
Official implementation for "RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers" (ICML 2025)
cogvideox diffusion diffusion-models diffusion-transformer dit extrapolation generative-model hunyuan-video long-video-generation position-embedding rope video-generation
Last synced: 01 Jul 2025
https://github.com/PKU-YuanGroup/MagicTime
MagicTime: Time-lapse Video Generation Models as Metamorphic Simulators
diffusion-models long-video-generation metamorphic-video-generation open-sora-plan text-to-video time-lapse time-lapse-dataset video-generation
Last synced: 11 Apr 2025
https://github.com/pku-yuangroup/magictime
MagicTime: Time-lapse Video Generation Models as Metamorphic Simulators
diffusion-models long-video-generation metamorphic-video-generation open-sora-plan text-to-video time-lapse time-lapse-dataset video-generation
Last synced: 04 Apr 2025
https://github.com/artokun/comfyui-mcp
Local-first, agent-native control plane for ComfyUI — MCP server + sidebar agent that generates images, video & audio, authors and runs workflows, and edits your live graph in natural language on ANY LLM (Claude, ChatGPT, Gemini, offline Ollama, or any hosted model). 178 tools, 36 AI skills, 55 installer packs. Local, LAN, VPS, or Comfy Cloud.
agent-skills ai-agent claude-code claude-plugin comfyui comfyui-extension comfyui-mcp comfyui-mcp-server flux image-generation local-first mcp mcp-server model-context-protocol offline ollama self-hosted stable-diffusion video-generation wan
Last synced: 27 Aug 2026
https://github.com/Kobaayyy/Awesome-CVPR2026-CVPR2025-ICCV2025-CVPR2024-ECCV2024-AIGC
A Collection of Papers and Codes for CVPR2026/CVPR2025/ICCV2025/CVPR2024/ECCV2024 AIGC
3d-generation aigc c-v-p-r cvpr cvpr2024 cvpr2025 cvpr2026 diffusion-models e-c-c-v eccv eccv2024 gan-models generative-ai iccv iccv2025 image-editing image-generation multi-modal-large-language-model video-editing video-generation
Last synced: 28 Feb 2026
https://github.com/simchowitzlabpublic/nano-world-model
A Minimalist, Batteries-included Repository for Advancing World Model Science.
diffusion-forcing diffusion-models model-predictive-control nano planning robot-manipulation streaming-video video-generation world-model
Last synced: 14 Sep 2026
https://github.com/Vchitect/VBench
[CVPR2024 Highlight] VBench - We Evaluate Video Generation
aigc benchmark dataset evaluation-kit gen-ai stable-diffusion text-to-video video-generation
Last synced: 30 Jul 2025
https://github.com/lucidrains/magvit2-pytorch
Implementation of MagViT2 Tokenizer in Pytorch
artificial-intelligence attention-mechanisms deep-learning finite-scalar-quantization transformers video-generation
Last synced: 14 May 2025
https://github.com/G-U-N/AnimateLCM
[SIGGRAPH ASIA 2024 TCS] AnimateLCM: Computation-Efficient Personalized Style Video Generation without Personalized Video Data
animatelcm consistency-models deep-learning fast-sampling video video-generation
Last synced: 28 Mar 2025
https://github.com/Kobaayyy/Awesome-CVPR2025-ICCV2025-CVPR2024-ECCV2024-AIGC
A Collection of Papers and Codes for CVPR2025/ICCV2025/CVPR2024/ECCV2024 AIGC
3d-generation aigc awesome c-v-p-r cvpr cvpr2024 cvpr2025 diffusion-models e-c-c-v eccv eccv2024 gan-models generative-ai iccv2025 image-editing image-generation multi-modal-large-language-model video-editing video-generation
Last synced: 03 Jul 2025
https://github.com/Orkas-AI/Orkas-VideoStudio
Turn your coding agent into a video studio: describe a video in plain language, and your agent writes the timeline and produces the file.
ai-agents ai-video cli coding-agents ffmpeg generative-ai hyperframes local-first mcp-server motion-graphics open-source subtitles text-to-video tts typescript video-automation video-composition video-editing video-generation whisper-cpp
Last synced: 18 Aug 2026
https://github.com/jaketae/storyteller
Multimodal AI Story Teller, built with Stable Diffusion, GPT, and neural text-to-speech
ddpm diffusion-models gpt image-generation natural-language-generation pytorch stable-diffusion text-to-image text-to-speech text-to-video video-generation
Last synced: 05 Apr 2025
https://github.com/vchitect/venhancer
Official codes of VEnhancer: Generative Space-Time Enhancement for Video Generation
aigc-enhancement diffusion-models frame-interpolation text-to-video video-enhancement video-generation video-super-resolution video-to-video
Last synced: 09 Apr 2025
https://github.com/Vchitect/VEnhancer
Official codes of VEnhancer: Generative Space-Time Enhancement for Video Generation
aigc-enhancement diffusion-models frame-interpolation text-to-video video-enhancement video-generation video-super-resolution video-to-video
Last synced: 28 Mar 2025
https://github.com/flymin/magicdrivedit
Official implementation of the paper “MagicDriveDiT: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control”
autonomous-driving diffusion-models multi-view video-generation
Last synced: 16 May 2025
https://github.com/samuraigpt/text-to-video-ai
Generate video from text using AI
ai-video-generator artificial-intelligence image-to-video image-to-video-generation sora-video-ai stable-diffusion text-to-image text-to-video text-to-video-generation video-diffusion video-diffusion-model video-editing video-generation
Last synced: 15 May 2025
https://github.com/cure-lab/MagicDrive
[ICLR24] Official implementation of the paper “MagicDrive: Street View Generation with Diverse 3D Geometry Control”
autonomous-vehicles deep-learning diffusion-models image-generation pytorch video-generation
Last synced: 20 Mar 2025
https://github.com/TianxingWu/FreeInit
[ECCV 2024] FreeInit: Bridging Initialization Gap in Video Diffusion Models
aigc text-to-video video-diffusion-model video-generation
Last synced: 28 Mar 2025
https://github.com/zhen-dong/magic-me
Codes for ID-Specific Video Customized Diffusion
diffusion-models image-animation image-editting personalized-generation text-to-video video-diffusion video-editing video-generation
Last synced: 06 Sep 2025
https://github.com/nihaomiao/CVPR23_LFDM
The pytorch implementation of our CVPR 2023 paper "Conditional Image-to-Video Generation with Latent Flow Diffusion Models"
cvpr2023 diffusion-models image-animation image-to-video latent-diffusion optical-flow video-generation video-prediction
Last synced: 28 Mar 2025
https://videoverses.github.io/videotuna/
Let's finetune video generation models!
ai aigc content-production fine-tuning-diffusion text-to-video video-generation visual-art
Last synced: 03 May 2025
https://github.com/wendell0218/Awesome-RL-for-Video-Generation
A curated list of papers on reinforcement learning for video generation
dpo grpo ppo reinforcement-learning reward-model video-generation
Last synced: 23 Aug 2026
https://github.com/EvoLinkAI/GPT-Image-2-Seedance2-Workflow
GPT-image-2 and seedance2 workflows and prompt templates to produce high-quality AI videos.
ai-video automation chatgpt content-creation creative-workflow generative-ai gpt-image-2 gpt-image-2-api gpt-image-2-prompts image-to-video media-pipeline multimodal-ai prompt-engineering seedance-2 seedance2 seedance2-api text-to-video video-generation visual-ai workflow-automation
Last synced: 18 Jun 2026