An open API service indexing awesome lists of open source software.

ai-game-devtools

Your AI Game Dev Hub. The ultimate resource hub for AI-powered game development tools. Discover cutting-edge LLMs, World Model, Agent, Code, Image, Texture, Shader, 3D Model, Animation, Video, Audio, Music, Singing Voice and Analytics. 🔥
https://github.com/Yuan-ManX/ai-game-devtools

Last synced: about 5 hours ago
JSON representation

  • Project List

    • <span id="tool">LLM (LLM & Tool)</span>

      • CogVLM - source visual language foundation model. |[arXiv](https://arxiv.org/abs/2311.03079) | | Tool |
      • Dora
      • GPT-4o - 4o (“o” for “omni”) is a step towards much more natural human-computer interaction—it accepts as input any combination of text, audio, image, and video and generates any combination of text, audio, and image outputs. | | | Tool |
      • Grok-1 - of-Experts model, Grok-1. | | | Tool |
      • HuggingChat
      • Mixtral 8x7B - of-Experts. |[arXiv](https://arxiv.org/abs/2401.04088) | | Tool |
      • Moshi
      • Nemotron-4 - billion-parameter large multilingual language model trained on 8 trillion text tokens. |[arXiv](https://arxiv.org/abs/2402.16819) | | Tool |
      • Pi
      • ShareGPT4V - Modal Models with Better Captions. | | | Tool |
      • NovelAI
      • AgentGPT
      • AICommand
      • AIOS
      • Assistant CLI
      • BabyAGI - powered task management system. | | | Tool |
      • 👶🤖🖥️ BabyAGI UI
      • baichuan-7B - scale 7B pretraining language model developed by Baichuan. | | | Tool |
      • Baichuan-13B
      • Baichuan 2
      • Bisheng
      • Character-LLM - Playing. |[arXiv](https://arxiv.org/abs/2310.10158) | | Tool |
      • ChatGPT-API-unity
      • ChatGPTForUnity
      • ChatRWKV
      • ChatYuan
      • Chinese-LLaMA-Alpaca-3 - 3 LLMs) developed from Meta Llama 3. | | | Tool |
      • Chrome-GPT
      • CoreNet
      • DBRX
      • DCLM
      • DemoGPT - AI App Generator with the Power of Llama 2 | | | Tool |
      • Design2Code - End Engineering | | | Tool |
      • Devika
      • Devon - source pair programmer. | | | Tool |
      • Flowise
      • Gemma - of-the art open models built from research and technology used to create Google Gemini models. | | | Tool |
      • gemma.cpp
      • GLM-4 - 4-9B is the open-source version of the latest generation of pre-trained models in the GLM-4 series launched by Zhipu AI. | | | Tool |
      • GPT4All
      • GPTScript
      • Hugging Face API Unity Integration - to-use integration for the Hugging Face Inference API, allowing developers to access and use Hugging Face AI models within their Unity projects. | | Unity | Tool |
      • ImageBind
      • Index-1.9B
      • InteractML-Unity
      • InternLM - sourced a 7 billion parameter base model, a chat model tailored for practical scenarios and the training system. |[arXiv](https://arxiv.org/abs/2403.17297) | | Tool |
      • Jan
      • Lamini - tuning on their own data. | | | Tool |
      • LaMini-LM - LM is a collection of small-sized, efficient language models distilled from ChatGPT and trained on a large-scale dataset of 2.58M instructions. | | | Tool |
      • LaVague
      • Lemur
      • Lepton AI
      • Lit-LLaMA - Adapter fine-tuning, pre-training. | | | Tool |
      • llama2-webui
      • Llama 3
      • Llama 3.1
      • LLaSM
      • LLM Answer Engine - Inspired Answer Engine Using Next.js, Groq, Mixtral, Langchain, OpenAI, Brave & Serper. | | | Tool |
      • llm.c
      • LLMUnity
      • LLocalSearch
      • LogicGamesSolver
      • Large World Model (LWM) - purpose large-context multimodal autoregressive model. |[arXiv](https://arxiv.org/abs/2402.08268) | | Tool |
      • Lumina-T2X - T2X is a unified framework for Text to Any Modality Generation. |[arXiv](https://arxiv.org/abs/2405.05945) | | Tool |
      • MetaGPT - Agent Framework | | | Tool |
      • MiniCPM-2B - side LLM outperforms Llama2-13B. | | | Tool |
      • MiniGPT-4 - language Understanding with Advanced Large Language Models. |[arXiv](https://arxiv.org/abs/2304.10592) | | Tool |
      • MiniGPT-5 - and-Language Generation via Generative Vokens. |[arXiv](https://arxiv.org/abs/2310.02239) | | Tool |
      • MLC LLM
      • MobiLlama
      • mPLUG-Owl🦉
      • NExT-GPT - to-Any Multimodal Large Language Model. | | | Tool |
      • OLMo
      • OneLLM
      • Open-Assistant - based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so. | | | Tool |
      • Orion-14B - 14B is a family of models includes a 14B foundation LLM, and a series of models. |[arXiv](https://arxiv.org/abs/2401.12246) | | Tool |
      • Panda - 7B, -13B, -33B, -65B for continuous pre-training in the Chinese field. | | | Tool |
      • Perplexica - powered search engine. | | | Tool |
      • RepoAgent - Source project driven by Large Language Models(LLMs) that aims to provide an intelligent way to document projects. |[arXiv](https://arxiv.org/abs/2402.16667) | | Tool |
      • Sanity AI Engine
      • SearchGPT
      • Skywork - trained on 3.2TB of high-quality multilingual (mainly Chinese and English) and code data. | | | Tool |
      • StableLM
      • Stanford Alpaca - following LLaMA Model. | | | LLM |
      • Text generation web UI - J, OPT, and GALACTICA. | | | Tool |
      • TinyChatEngine - Device LLM Inference Library. | | | Tool |
      • ToolBench
      • Unity ChatGPT
      • Unreal Engine 5 Llama LoRA - of-concept project that showcases the potential for using small, locally trainable LLMs to create next-generation documentation tools. | | Unreal Engine | Tool |
      • UnrealGPT
      • WebGPT
      • Web3-GPT
      • WordGPT
      • Yi
      • 01 Project - source language model computer. | | | Tool |
      • AI-Writer - trained generative model. | | | Writer |
      • Notebook.ai
      • Novel - style WYSIWYG editor with AI-powered autocompletions. | | | Writer |
      • AI Scientist - Ended Scientific Discovery. |[arXiv](https://arxiv.org/abs/2408.06292) | | Tool |
      • LongWriter
      • Moshi - text foundation model for real time dialogue. | | | Tool |
      • DeepSeek-V3 - V3 is a strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token. |[arXiv](https://arxiv.org/abs/2412.19437) | | LLM |
      • Cosmos
      • MiniMax-01 - 01: Scaling Foundation Models with Lightning Attention. |[arXiv](https://arxiv.org/abs/2501.08313) | | LLM |
      • SkyThought - T1: Train your own O1 preview model within $450. | | | LLM |
      • DeepSeek-R1 - R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. | | | LLM |
      • Janus
      • s1 - time scaling. |[arXiv](https://arxiv.org/abs/2501.19393) | | LLM |
      • Open Deep Research - powered research assistant that performs iterative, deep research on any topic by combining search engines, web scraping, and large language models. | | | LLM |
      • LangChain
      • OpenDevin
      • Gemini
      • SimpleOllamaUnity
      • GLM-4.5 - 4.5: An open-source large language model designed for intelligent agents by Z.ai. | | | LLM |
      • gpt-oss - oss-120b and gpt-oss-20b are two open-weight language models by OpenAI. | | | LLM |
      • Kimi K2 - of-the-art mixture-of-experts (MoE) language model with 32 billion activated parameters and 1 trillion total parameters. | | | LLM |
      • Qwen3
      • Seed-OSS - OSS is a series of open-source large language models developed by ByteDance's Seed Team, designed for powerful long-context, reasoning, agent and general capabilities, and versatile developer-friendly features. | | | LLM |
      • LongCat-Flash - Flash is a powerful and efficient language model with 560 billion total parameters, featuring an innovative Mixture-of-Experts (MoE) architecture. The model incorporates a dynamic computation mechanism that activates 18.6B∼31.3B parameters (averaging∼27B) based on contextual demands, optimizing both computational efficiency and performance. | | | LLM |
      • Hunyuan-MT - MT comprises a translation model, Hunyuan-MT-7B, and an ensemble model, Hunyuan-MT-Chimera. The translation model is used to translate source text into the target language, while the ensemble model integrates multiple translation outputs to produce a higher-quality result. | | | LLM |
      • MOSS - source tool-augmented conversational language model from Fudan University. | | | Tool |
      • Auto-GPT - source attempt to make GPT-4 fully autonomous. | | | Tool |
      • Qwen1.5
      • Qwen-7B - 7B (通义千问-7B) chat & pretrained large language model proposed by Alibaba Cloud. | | | LLM |
      • Qwen2
      • Mistral 7B
      • Mistral Large - edge text generation model. It reaches top-tier reasoning capabilities. | | | Tool |
      • Mixtral 8x7B - of-Experts. |[arXiv](https://arxiv.org/abs/2401.04088) | | Tool |
      • Auferet - style RPGs, with persistent memory of your story and your own uploaded lore. | | | Writer |
      • Unity-MCP - source MCP server connecting AI agents to the Unity Editor and runtime, with 100+ built-in tools. | | Unity | Tool |
      • Godot-MCP - source MCP server connecting AI agents to the Godot Editor and runtime (Godot 4.x, C#). | | Godot | Tool |
      • Unreal-MCP - source MCP server connecting AI agents to Unreal Engine 5.7, editor and runtime (C++ plugin + .NET sidecar). | | Unreal Engine | Tool |
      • GameDev-MCP-Server - source, engine-agnostic MCP server shared by Unity-MCP, Godot-MCP, and Unreal-MCP. | | Unity/Godot/Unreal Engine | Tool |
      • MCP-Plugin-dotnet - source .NET library/SDK that turns any .NET application into an MCP server. | | | Tool |
      • ReflectorNet - source .NET reflection toolkit for AI-driven scenarios. | | | Tool |
      • Grok-1 - of-Experts model, Grok-1. | | | Tool |
      • InteractML-Unreal Engine
      • LangFlow - flow to provide an effortless way to experiment and prototype flows. | | | Tool |
      • OmniLMM - modal models for strong performance and efficient deployment. | | | Tool |
      • Unity OpenAI-API Integration - 3 language model and ChatGPT API into a Unity project. | | Unity | Tool |
    • <span id="tool">Tool (AI LLM)</span>

  • <span id="animation">Animation</span>

    • <span id="tool">LLM (LLM & Tool)</span>

      • Deforum
      • FreeInit
      • ID-Animator - Shot Identity-Preserving Human Video Generation. |[arXiv](https://arxiv.org/abs/2404.15275) | | Animation |
      • NUWA-XL
      • NUWA-Infinity - Infinity is a multimodal generative model that is designed to generate high-quality images and videos from given text, image or video input. | | | Animation |
      • PIA - and-Play Modules in Text-to-Image Models. |[arXiv](https://arxiv.org/abs/2312.13964) | | Animation |
      • Stable Animation - to-animation tool for developers. | | | Animation |
      • Wonder Studio - action scene. | | | Animation |
      • Animate Anyone - to-Video Synthesis for Character Animation. |[arXiv](https://arxiv.org/abs/2311.17117) | | Animation |
      • AnimateAnything - Grained Open Domain Image Animation with Motion Guidance. |[arXiv](https://arxiv.org/abs/2311.12886) | | Animation |
      • AnimateLCM
      • AnimationGPT
      • DreaMoving
      • FaceFusion
      • GeneFace - Fidelity Audio-Driven 3D Talking Face Synthesis. |[arXiv](https://arxiv.org/abs/2301.13430) | | Animation |
      • MagicAnimate
      • SadTalker-Video-Lip-Sync
      • Wav2Lip - syncing Videos In The Wild. |[arXiv](https://arxiv.org/abs/2008.10010) | | Animation |
      • DrawingSpinUp
      • Animate-X - X: Universal Character Image Animation with Enhanced Motion Representation. |[arXiv](https://arxiv.org/abs/2410.10306) | | Animation |
      • Omni Animation
      • AnimateZero - Shot Image Animators. |[arXiv](https://arxiv.org/abs/2312.03793) | | Animation |
      • Index-AniSora - AniSora is the most powerful open-source animated video generation model. It enables one-click creation of video shots across diverse anime styles including series episodes, Chinese original animations, manga adaptations, VTuber content, anime PVs, mad-style parodies(鬼畜动画), and more! |[arXiv](https://arxiv.org/abs/2412.10255) | | Animation |
      • ToonComposer - Keyframing. |[arXiv](https://arxiv.org/abs/2508.10881) | | Animation |
      • AnimateDiff - to-Image Diffusion Models without Specific Tuning. |[arXiv](https://arxiv.org/abs/2307.04725) | | Animation |
      • SadTalker - Driven Single Image Talking Face Animation. |[arXiv](https://arxiv.org/abs/2211.12194) | | Animation |
      • TaleCrafter
      • HY-Motion 1.0 - Motion 1.0: Scaling Flow Matching Models for Text-To-Motion Generation. |[arXiv](https://arxiv.org/abs/2512.23464) | | Animation |
      • ToonCrafter
    • <span id="tool">Tool (AI LLM)</span>

  • <span id="audio">Audio</span>

    • <span id="tool">LLM (LLM & Tool)</span>

      • Audiobox
      • AudioLDM - to-Audio Generation with Latent Diffusion Models. |[arXiv](https://arxiv.org/abs/2301.12503) | | Audio |
      • MAGNeT - Autoregressive Transformer. | | | Audio |
      • Make-An-Audio - To-Audio Generation with Prompt-Enhanced Diffusion Models. |[arXiv](https://arxiv.org/abs/2301.12661) | | Audio |
      • OptimizerAI
      • SoundStorm
      • Stable Audio - Conditioned Latent Audio Diffusion. | | | Audio |
      • Stable Audio Open - length (up to 47s) stereo audio at 44.1kHz from text prompts. | | | Audio |
      • FoleyCrafter
      • SyncFusion - synchronized Video-to-Audio Foley Synthesis. |[arXiv](https://arxiv.org/abs/2310.15247) | | Audio |
      • AcademiCodec
      • Amphion - Source Audio, Music, and Speech Generation Toolkit. |[arXiv](https://arxiv.org/abs/2312.09911) | | Audio |
      • ArchiSound
      • AudioEditing - Shot Unsupervised and Text-Based Audio Editing Using DDPM Inversion. |[arXiv](https://arxiv.org/abs/2402.10009) | | Audio |
      • Audiogen Codec
      • AudioGPT
      • AudioLCM - to-Audio Generation with Latent Consistency Models. |[arXiv](https://arxiv.org/abs/2406.00356v1) | | Audio |
      • AudioLDM 2 - supervised Pretraining. |[arXiv](https://arxiv.org/abs/2308.05734) | | Audio |
      • Auffusion - to-Audio Generation. |[arXiv](https://arxiv.org/abs/2401.01044) | | Audio |
      • CTAG - to-Audio Generation via Synthesizer Programming. | | | Audio |
      • Make-An-Audio 3 - based Large Diffusion Transformers. |[arXiv](https://arxiv.org/abs/2305.18474) | | Audio |
      • NeuralSound - based Modal Sound Synthesis with Acoustic Transfer. |[arXiv](https://arxiv.org/abs/2108.07425) | | Audio |
      • Qwen2-Audio - Audio chat & pretrained large audio language model proposed by Alibaba Cloud. |[arXiv](https://arxiv.org/abs/2407.10759) | | Audio |
      • SEE-2-SOUND - Shot Spatial Environment-to-Spatial Sound. |[arXiv](https://arxiv.org/abs/2406.06612) | | Audio |
      • TANGO - to-Audio Generation using Instruction Tuned LLM and Latent Diffusion Model. | | | Audio |