ai-game-devtools
Your AI Game Dev Hub. The ultimate resource hub for AI-powered game development tools. Discover cutting-edge LLMs, World Model, Agent, Code, Image, Texture, Shader, 3D Model, Animation, Video, Audio, Music, Singing Voice and Analytics. 🔥
https://github.com/Yuan-ManX/ai-game-devtools
Last synced: about 7 hours ago
JSON representation
-
Project List
-
<span id="tool">LLM (LLM & Tool)</span>
- CogVLM - source visual language foundation model. |[arXiv](https://arxiv.org/abs/2311.03079) | | Tool |
- Dora
- GPT-4o - 4o (“o” for “omni”) is a step towards much more natural human-computer interaction—it accepts as input any combination of text, audio, image, and video and generates any combination of text, audio, and image outputs. | | | Tool |
- Grok-1 - of-Experts model, Grok-1. | | | Tool |
- HuggingChat
- Mixtral 8x7B - of-Experts. |[arXiv](https://arxiv.org/abs/2401.04088) | | Tool |
- Moshi
- Nemotron-4 - billion-parameter large multilingual language model trained on 8 trillion text tokens. |[arXiv](https://arxiv.org/abs/2402.16819) | | Tool |
- Pi
- ShareGPT4V - Modal Models with Better Captions. | | | Tool |
- NovelAI
- AgentGPT
- AICommand
- AIOS
- Assistant CLI
- BabyAGI - powered task management system. | | | Tool |
- 👶🤖🖥️ BabyAGI UI
- baichuan-7B - scale 7B pretraining language model developed by Baichuan. | | | Tool |
- Baichuan-13B
- Baichuan 2
- Bisheng
- Character-LLM - Playing. |[arXiv](https://arxiv.org/abs/2310.10158) | | Tool |
- ChatGPT-API-unity
- ChatGPTForUnity
- ChatRWKV
- ChatYuan
- Chinese-LLaMA-Alpaca-3 - 3 LLMs) developed from Meta Llama 3. | | | Tool |
- Chrome-GPT
- CoreNet
- DBRX
- DCLM
- DemoGPT - AI App Generator with the Power of Llama 2 | | | Tool |
- Design2Code - End Engineering | | | Tool |
- Devika
- Devon - source pair programmer. | | | Tool |
- Flowise
- Gemma - of-the art open models built from research and technology used to create Google Gemini models. | | | Tool |
- gemma.cpp
- GLM-4 - 4-9B is the open-source version of the latest generation of pre-trained models in the GLM-4 series launched by Zhipu AI. | | | Tool |
- GPT4All
- GPTScript
- Hugging Face API Unity Integration - to-use integration for the Hugging Face Inference API, allowing developers to access and use Hugging Face AI models within their Unity projects. | | Unity | Tool |
- ImageBind
- Index-1.9B
- InteractML-Unity
- InternLM - sourced a 7 billion parameter base model, a chat model tailored for practical scenarios and the training system. |[arXiv](https://arxiv.org/abs/2403.17297) | | Tool |
- Jan
- Lamini - tuning on their own data. | | | Tool |
- LaMini-LM - LM is a collection of small-sized, efficient language models distilled from ChatGPT and trained on a large-scale dataset of 2.58M instructions. | | | Tool |
- LaVague
- Lemur
- Lepton AI
- Lit-LLaMA - Adapter fine-tuning, pre-training. | | | Tool |
- llama2-webui
- Llama 3
- Llama 3.1
- LLaSM
- LLM Answer Engine - Inspired Answer Engine Using Next.js, Groq, Mixtral, Langchain, OpenAI, Brave & Serper. | | | Tool |
- llm.c
- LLMUnity
- LLocalSearch
- LogicGamesSolver
- Large World Model (LWM) - purpose large-context multimodal autoregressive model. |[arXiv](https://arxiv.org/abs/2402.08268) | | Tool |
- Lumina-T2X - T2X is a unified framework for Text to Any Modality Generation. |[arXiv](https://arxiv.org/abs/2405.05945) | | Tool |
- MetaGPT - Agent Framework | | | Tool |
- MiniCPM-2B - side LLM outperforms Llama2-13B. | | | Tool |
- MiniGPT-4 - language Understanding with Advanced Large Language Models. |[arXiv](https://arxiv.org/abs/2304.10592) | | Tool |
- MiniGPT-5 - and-Language Generation via Generative Vokens. |[arXiv](https://arxiv.org/abs/2310.02239) | | Tool |
- MLC LLM
- MobiLlama
- mPLUG-Owl🦉
- NExT-GPT - to-Any Multimodal Large Language Model. | | | Tool |
- OLMo
- OneLLM
- Open-Assistant - based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so. | | | Tool |
- Orion-14B - 14B is a family of models includes a 14B foundation LLM, and a series of models. |[arXiv](https://arxiv.org/abs/2401.12246) | | Tool |
- Panda - 7B, -13B, -33B, -65B for continuous pre-training in the Chinese field. | | | Tool |
- Perplexica - powered search engine. | | | Tool |
- RepoAgent - Source project driven by Large Language Models(LLMs) that aims to provide an intelligent way to document projects. |[arXiv](https://arxiv.org/abs/2402.16667) | | Tool |
- Sanity AI Engine
- SearchGPT
- Skywork - trained on 3.2TB of high-quality multilingual (mainly Chinese and English) and code data. | | | Tool |
- StableLM
- Stanford Alpaca - following LLaMA Model. | | | LLM |
- Text generation web UI - J, OPT, and GALACTICA. | | | Tool |
- TinyChatEngine - Device LLM Inference Library. | | | Tool |
- ToolBench
- Unity ChatGPT
- Unreal Engine 5 Llama LoRA - of-concept project that showcases the potential for using small, locally trainable LLMs to create next-generation documentation tools. | | Unreal Engine | Tool |
- UnrealGPT
- WebGPT
- Web3-GPT
- WordGPT
- Yi
- 01 Project - source language model computer. | | | Tool |
- AI-Writer - trained generative model. | | | Writer |
- Notebook.ai
- Novel - style WYSIWYG editor with AI-powered autocompletions. | | | Writer |
- AI Scientist - Ended Scientific Discovery. |[arXiv](https://arxiv.org/abs/2408.06292) | | Tool |
- LongWriter
- Moshi - text foundation model for real time dialogue. | | | Tool |
- DeepSeek-V3 - V3 is a strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token. |[arXiv](https://arxiv.org/abs/2412.19437) | | LLM |
- Cosmos
- MiniMax-01 - 01: Scaling Foundation Models with Lightning Attention. |[arXiv](https://arxiv.org/abs/2501.08313) | | LLM |
- SkyThought - T1: Train your own O1 preview model within $450. | | | LLM |
- DeepSeek-R1 - R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. | | | LLM |
- Janus
- s1 - time scaling. |[arXiv](https://arxiv.org/abs/2501.19393) | | LLM |
- Open Deep Research - powered research assistant that performs iterative, deep research on any topic by combining search engines, web scraping, and large language models. | | | LLM |
- LangChain
- OpenDevin
- Gemini
- SimpleOllamaUnity
- GLM-4.5 - 4.5: An open-source large language model designed for intelligent agents by Z.ai. | | | LLM |
- gpt-oss - oss-120b and gpt-oss-20b are two open-weight language models by OpenAI. | | | LLM |
- Kimi K2 - of-the-art mixture-of-experts (MoE) language model with 32 billion activated parameters and 1 trillion total parameters. | | | LLM |
- Qwen3
- Seed-OSS - OSS is a series of open-source large language models developed by ByteDance's Seed Team, designed for powerful long-context, reasoning, agent and general capabilities, and versatile developer-friendly features. | | | LLM |
- LongCat-Flash - Flash is a powerful and efficient language model with 560 billion total parameters, featuring an innovative Mixture-of-Experts (MoE) architecture. The model incorporates a dynamic computation mechanism that activates 18.6B∼31.3B parameters (averaging∼27B) based on contextual demands, optimizing both computational efficiency and performance. | | | LLM |
- Hunyuan-MT - MT comprises a translation model, Hunyuan-MT-7B, and an ensemble model, Hunyuan-MT-Chimera. The translation model is used to translate source text into the target language, while the ensemble model integrates multiple translation outputs to produce a higher-quality result. | | | LLM |
- MOSS - source tool-augmented conversational language model from Fudan University. | | | Tool |
- Auto-GPT - source attempt to make GPT-4 fully autonomous. | | | Tool |
- Qwen1.5
- Qwen-7B - 7B (通义千问-7B) chat & pretrained large language model proposed by Alibaba Cloud. | | | LLM |
- Qwen2
- Mistral 7B
- Mistral Large - edge text generation model. It reaches top-tier reasoning capabilities. | | | Tool |
- Mixtral 8x7B - of-Experts. |[arXiv](https://arxiv.org/abs/2401.04088) | | Tool |
- Auferet - style RPGs, with persistent memory of your story and your own uploaded lore. | | | Writer |
- Unity-MCP - source MCP server connecting AI agents to the Unity Editor and runtime, with 100+ built-in tools. | | Unity | Tool |
- Godot-MCP - source MCP server connecting AI agents to the Godot Editor and runtime (Godot 4.x, C#). | | Godot | Tool |
- Unreal-MCP - source MCP server connecting AI agents to Unreal Engine 5.7, editor and runtime (C++ plugin + .NET sidecar). | | Unreal Engine | Tool |
- GameDev-MCP-Server - source, engine-agnostic MCP server shared by Unity-MCP, Godot-MCP, and Unreal-MCP. | | Unity/Godot/Unreal Engine | Tool |
- MCP-Plugin-dotnet - source .NET library/SDK that turns any .NET application into an MCP server. | | | Tool |
- ReflectorNet - source .NET reflection toolkit for AI-driven scenarios. | | | Tool |
- Grok-1 - of-Experts model, Grok-1. | | | Tool |
- InteractML-Unreal Engine
- LangFlow - flow to provide an effortless way to experiment and prototype flows. | | | Tool |
- OmniLMM - modal models for strong performance and efficient deployment. | | | Tool |
- Unity OpenAI-API Integration - 3 language model and ChatGPT API into a Unity project. | | Unity | Tool |
-
<span id="tool">Tool (AI LLM)</span>
- Mistral 7B
- InteractML-Unreal Engine
- Unity OpenAI-API Integration - 3 language model and ChatGPT API into a Unity project. | | Unity | Tool |
- Mistral Large - edge text generation model. It reaches top-tier reasoning capabilities. | | | Tool |
-
-
<span id="animation">Animation</span>
-
<span id="tool">LLM (LLM & Tool)</span>
- Deforum
- FreeInit
- ID-Animator - Shot Identity-Preserving Human Video Generation. |[arXiv](https://arxiv.org/abs/2404.15275) | | Animation |
- NUWA-XL
- NUWA-Infinity - Infinity is a multimodal generative model that is designed to generate high-quality images and videos from given text, image or video input. | | | Animation |
- PIA - and-Play Modules in Text-to-Image Models. |[arXiv](https://arxiv.org/abs/2312.13964) | | Animation |
- Stable Animation - to-animation tool for developers. | | | Animation |
- Wonder Studio - action scene. | | | Animation |
- Animate Anyone - to-Video Synthesis for Character Animation. |[arXiv](https://arxiv.org/abs/2311.17117) | | Animation |
- AnimateAnything - Grained Open Domain Image Animation with Motion Guidance. |[arXiv](https://arxiv.org/abs/2311.12886) | | Animation |
- AnimateLCM
- AnimationGPT
- DreaMoving
- FaceFusion
- GeneFace - Fidelity Audio-Driven 3D Talking Face Synthesis. |[arXiv](https://arxiv.org/abs/2301.13430) | | Animation |
- MagicAnimate
- SadTalker-Video-Lip-Sync
- Wav2Lip - syncing Videos In The Wild. |[arXiv](https://arxiv.org/abs/2008.10010) | | Animation |
- DrawingSpinUp
- Animate-X - X: Universal Character Image Animation with Enhanced Motion Representation. |[arXiv](https://arxiv.org/abs/2410.10306) | | Animation |
- Omni Animation
- AnimateZero - Shot Image Animators. |[arXiv](https://arxiv.org/abs/2312.03793) | | Animation |
- Index-AniSora - AniSora is the most powerful open-source animated video generation model. It enables one-click creation of video shots across diverse anime styles including series episodes, Chinese original animations, manga adaptations, VTuber content, anime PVs, mad-style parodies(鬼畜动画), and more! |[arXiv](https://arxiv.org/abs/2412.10255) | | Animation |
- ToonComposer - Keyframing. |[arXiv](https://arxiv.org/abs/2508.10881) | | Animation |
- AnimateDiff - to-Image Diffusion Models without Specific Tuning. |[arXiv](https://arxiv.org/abs/2307.04725) | | Animation |
- SadTalker - Driven Single Image Talking Face Animation. |[arXiv](https://arxiv.org/abs/2211.12194) | | Animation |
- TaleCrafter
- HY-Motion 1.0 - Motion 1.0: Scaling Flow Matching Models for Text-To-Motion Generation. |[arXiv](https://arxiv.org/abs/2512.23464) | | Animation |
- ToonCrafter
-
<span id="tool">Tool (AI LLM)</span>
- Omni Animation
- AnimateZero - Shot Image Animators. |[arXiv](https://arxiv.org/abs/2312.03793) | | Animation |
-
-
<span id="audio">Audio</span>
-
<span id="tool">LLM (LLM & Tool)</span>
- Audiobox
- AudioLDM - to-Audio Generation with Latent Diffusion Models. |[arXiv](https://arxiv.org/abs/2301.12503) | | Audio |
- MAGNeT - Autoregressive Transformer. | | | Audio |
- Make-An-Audio - To-Audio Generation with Prompt-Enhanced Diffusion Models. |[arXiv](https://arxiv.org/abs/2301.12661) | | Audio |
- OptimizerAI
- SoundStorm
- Stable Audio - Conditioned Latent Audio Diffusion. | | | Audio |
- Stable Audio Open - length (up to 47s) stereo audio at 44.1kHz from text prompts. | | | Audio |
- FoleyCrafter
- SyncFusion - synchronized Video-to-Audio Foley Synthesis. |[arXiv](https://arxiv.org/abs/2310.15247) | | Audio |
- AcademiCodec
- Amphion - Source Audio, Music, and Speech Generation Toolkit. |[arXiv](https://arxiv.org/abs/2312.09911) | | Audio |
- ArchiSound
- AudioEditing - Shot Unsupervised and Text-Based Audio Editing Using DDPM Inversion. |[arXiv](https://arxiv.org/abs/2402.10009) | | Audio |
- Audiogen Codec
- AudioGPT
- AudioLCM - to-Audio Generation with Latent Consistency Models. |[arXiv](https://arxiv.org/abs/2406.00356v1) | | Audio |
- AudioLDM 2 - supervised Pretraining. |[arXiv](https://arxiv.org/abs/2308.05734) | | Audio |
- Auffusion - to-Audio Generation. |[arXiv](https://arxiv.org/abs/2401.01044) | | Audio |
- CTAG - to-Audio Generation via Synthesizer Programming. | | | Audio |
- Make-An-Audio 3 - based Large Diffusion Transformers. |[arXiv](https://arxiv.org/abs/2305.18474) | | Audio |
- NeuralSound - based Modal Sound Synthesis with Acoustic Transfer. |[arXiv](https://arxiv.org/abs/2108.07425) | | Audio |
- Qwen2-Audio - Audio chat & pretrained large audio language model proposed by Alibaba Cloud. |[arXiv](https://arxiv.org/abs/2407.10759) | | Audio |
- SEE-2-SOUND - Shot Spatial Environment-to-Spatial Sound. |[arXiv](https://arxiv.org/abs/2406.06612) | | Audio |
- TANGO - to-Audio Generation using Instruction Tuned LLM and Latent Diffusion Model. | | | Audio |
-
Programming Languages
Categories
Project List
144
<span id="video">Video</span>
138
<span id="image">Image</span>
113
<span id="game">Game (World Model & Agent)</span>
82
<span id="model">3D Model</span>
73
<span id="speech">Speech</span>
57
<span id="avatar">Avatar</span>
40
<span id="audio">Audio</span>
33
<span id="code">Code</span>
32
<span id="visual">VLM (Visual)</span>
32
<span id="animation">Animation</span>
31
<span id="music">Music</span>
26
<span id="texture">Texture</span>
21
<span id="voice">Singing Voice</span>
4
<span id="shader">Shader</span>
1
<span id="visual">Visual</span>
1
<span id="game">Game (Agent)</span>
1
<span id="speech">Analytics</span>
1
Keywords
llm
45
ai
43
chatgpt
32
diffusion-models
26
gpt
24
large-language-models
24
text-to-speech
20
tts
20
openai
20
pytorch
19
python
19
deep-learning
18
agent
18
stable-diffusion
16
artificial-intelligence
16
gpt-4
15
chatbot
15
computer-vision
14
language-model
14
aigc
13
video-generation
12
machine-learning
11
rag
11
speech-synthesis
11
generative-ai
11
3d-generation
10
unity
10
llama
10
langchain
9
diffusion
9
audio-generation
8
speech
8
video
8
generative-model
8
multimodal
8
image-generation
7
chinese
7
code-generation
7
multi-modal
7
agents
6
instruction-tuning
6
vits
6
image-to-3d
6
voice-cloning
6
nextjs
6
nerf
6
transformers
6
voice-conversion
5
video-editing
5
pretrained-models
5