{"id":63306,"url":"https://github.com/Yuan-ManX/ai-game-devtools","name":"ai-game-devtools","description":"Your AI Game Dev Hub. The ultimate resource hub for AI-powered game development tools. Discover cutting-edge LLMs, World Model, Agent, Code, Image, Texture, Shader, 3D Model, Animation, Video, Audio, Music, Singing Voice and Analytics. 🔥","projects_count":807,"last_synced_at":"2026-09-14T23:00:18.223Z","repository":{"id":152524633,"uuid":"616764080","full_name":"Yuan-ManX/ai-game-devtools","owner":"Yuan-ManX","description":"Your AI Game Dev Hub. The ultimate resource hub for AI-powered game development tools. Discover cutting-edge LLMs, World Model, Agent, Code, Image, Texture, Shader, 3D Model, Animation, Video, Audio, Music, Singing Voice and Analytics. 🔥","archived":false,"fork":false,"pushed_at":"2026-07-21T09:21:18.000Z","size":7438,"stargazers_count":1316,"open_issues_count":11,"forks_count":126,"subscribers_count":46,"default_branch":"main","last_synced_at":"2026-08-26T04:08:06.312Z","etag":null,"topics":["ai-agents","ai-game-development","ai-game-engine","ai-platform","ai-toolkit","aigc","artificial-intelligence","awesome-list","deep-learning","game-ai","game-development","game-engine","mechine-learing","unity","world-models"],"latest_commit_sha":null,"homepage":"https://yuan-manx.github.io/ai-game-devtools/","language":"JavaScript","has_issues":true,"has_wiki":null,"has_pages":null,"mirror_url":null,"source_name":null,"license":"mit","status":null,"scm":"git","pull_requests_enabled":true,"icon_url":"https://github.com/Yuan-ManX.png","metadata":{"files":{"readme":"README.md","changelog":null,"contributing":null,"funding":null,"license":"LICENSE","code_of_conduct":null,"threat_model":null,"audit":null,"citation":null,"codeowners":null,"security":null,"support":null,"governance":null,"roadmap":null,"authors":null,"dei":null,"publiccode":null,"codemeta":null,"zenodo":null,"notice":null,"maintainers":null,"copyright":null,"agents":null,"dco":null,"cla":null}},"created_at":"2023-03-21T03:01:17.000Z","updated_at":"2026-08-25T20:11:19.000Z","dependencies_parsed_at":"2023-10-25T10:26:32.296Z","dependency_job_id":"a312772b-fdc9-459a-ba3b-79adf65c8188","html_url":"https://github.com/Yuan-ManX/ai-game-devtools","commit_stats":null,"previous_names":["yuan-manx/ai-game-devtools","yuan-manx/ai-game-development-tools"],"tags_count":0,"template":false,"template_full_name":null,"purl":"pkg:github/Yuan-ManX/ai-game-devtools","repository_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Yuan-ManX%2Fai-game-devtools","tags_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Yuan-ManX%2Fai-game-devtools/tags","releases_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Yuan-ManX%2Fai-game-devtools/releases","manifests_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Yuan-ManX%2Fai-game-devtools/manifests","owner_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners/Yuan-ManX","download_url":"https://codeload.github.com/Yuan-ManX/ai-game-devtools/tar.gz/refs/heads/main","sbom_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/Yuan-ManX%2Fai-game-devtools/sbom","scorecard":null,"host":{"name":"GitHub","url":"https://github.com","kind":"github","repositories_count":341189360,"owners_count":37327249,"icon_url":"https://github.com/github.png","version":null,"created_at":"2022-05-30T11:31:42.601Z","updated_at":"2026-08-22T15:14:58.755Z","status":"online","status_checked_at":"2026-09-14T02:00:06.290Z","response_time":176,"last_error":null,"robots_txt_status":"success","robots_txt_updated_at":"2025-07-24T06:49:26.215Z","robots_txt_url":"https://github.com/robots.txt","online":true,"can_crawl_api":true,"host_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub","repositories_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories","repository_names_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repository_names","owners_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners"}},"created_at":"2024-07-08T05:34:31.544Z","updated_at":"2026-09-14T23:00:18.224Z","primary_language":"Python","list_of_lists":false,"displayable":true,"categories":["\u003cspan id=\"avatar\"\u003eAvatar\u003c/span\u003e","\u003cspan id=\"model\"\u003e3D Model\u003c/span\u003e","\u003cspan id=\"image\"\u003eImage\u003c/span\u003e","\u003cspan id=\"video\"\u003eVideo\u003c/span\u003e","\u003cspan id=\"music\"\u003eMusic\u003c/span\u003e","Project List","\u003cspan id=\"speech\"\u003eSpeech\u003c/span\u003e","\u003cspan id=\"animation\"\u003eAnimation\u003c/span\u003e","\u003cspan id=\"audio\"\u003eAudio\u003c/span\u003e","\u003cspan id=\"code\"\u003eCode\u003c/span\u003e","\u003cspan id=\"game\"\u003eGame (World Model \u0026 Agent)\u003c/span\u003e","\u003cspan id=\"texture\"\u003eTexture\u003c/span\u003e","\u003cspan id=\"speech\"\u003eAnalytics\u003c/span\u003e","\u003cspan id=\"visual\"\u003eVLM (Visual)\u003c/span\u003e","\u003cspan id=\"voice\"\u003eSinging Voice\u003c/span\u003e","\u003cspan id=\"shader\"\u003eShader\u003c/span\u003e","\u003cspan id=\"game\"\u003eGame (Agent)\u003c/span\u003e","\u003cspan id=\"visual\"\u003eVisual\u003c/span\u003e"],"sub_categories":["\u003cspan id=\"tool\"\u003eLLM (LLM \u0026 Tool)\u003c/span\u003e","\u003cspan id=\"tool\"\u003eTool (AI LLM)\u003c/span\u003e"],"readme":"\u003cdiv align=\"center\"\u003e\n\n\u003cimg src=\"./assets/AI Game DevTools.png\" alt=\"AI Game DevTools\"\u003e\n\n# AI Game DevTools (AI-GDT) 🎮\n\n### Your AI Game Dev Hub. 🚀\n\nThe ultimate resource hub for AI-powered game development tools. Discover cutting-edge LLMs, World Model, Agent, Code, Image, Texture, Shader, 3D Model, Animation, Video, Audio, Music, Singing Voice and Analytics. 🔥\n\n![Version](https://img.shields.io/badge/version-1.0.0-blue)\n![License](https://img.shields.io/badge/license-MIT-green)\n![Stars](https://img.shields.io/github/stars/Yuan-ManX/ai-game-devtools?style=social)\n\n### [Website](https://yuan-manx.github.io/ai-game-devtools/) | [官方网站](https://yuan-manx.github.io/ai-game-devtools/)\n\n\u003c/div\u003e\n\n\n## Table of Contents\n\n* [LLM (LLM \u0026 Tool)](#tool)\n* [VLM (Visual)](#visual)\n* [Game (World Model \u0026 Agent)](#game)\n* [Code](#code)\n* [Image](#image)\n* [Texture](#texture)\n* [Shader](#shader)\n* [3D Model](#model)\n* [Avatar](#avatar)\n* [Animation](#animation)\n* [Video](#video)\n* [Audio](#audio)\n* [Music](#music)\n* [Singing Voice](#voice)\n* [Speech](#speech)\n* [Analytics](#analytics)\n\n\n## Project List\n\n###  \u003cspan id=\"tool\"\u003eLLM (LLM \u0026 Tool)\u003c/span\u003e\n\n| Source                                                                                      | Description                                                                                                                                                                                    |   Paper   |  Game Engine  |   Type   |\n| :------------------------------------------------------------------------------------------ | :--------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | :-----------: | :-----------: | :-------: |\n| [AgentGPT](https://github.com/reworkd/AgentGPT)                                                | 🤖 Assemble, configure, and deploy autonomous AI Agents in your browser.                                                                                                                      |          |              |   Tool   |\n| [AICommand](https://github.com/keijiro/AICommand)                                              | ChatGPT integration with Unity Editor.                                                                                                                                                         |          |     Unity    |   Tool   |\n| [AIOS](https://github.com/agiresearch/AIOS)                                                    | LLM Agent Operating System.                                                                                                                                                                    |          |              |   Tool   |\n| [AI Scientist](https://github.com/SakanaAI/AI-Scientist)                                       | The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery.                                                       |[arXiv](https://arxiv.org/abs/2408.06292)  |             |   Tool   |\n| [Assistant CLI](https://github.com/diciaup/assistant-cli)                                      | A comfortable CLI tool to use ChatGPT service🔥                                                                                                                                               |          |              |   Tool   |\n| [Auferet](https://auferet.com/) | An AI game master for solo text adventures and tabletop-style RPGs, with persistent memory of your story and your own uploaded lore. |  |  | Writer |\n| [Auto-GPT](https://github.com/Significant-Gravitas/Auto-GPT)                                   | An experimental open-source attempt to make GPT-4 fully autonomous.                                                                                                                            |           |             |   Tool   |\n| [BabyAGI](https://github.com/yoheinakajima/babyagi)                                            | This Python script is an example of an AI-powered task management system.                                                                                                                      |          |              |   Tool   |\n| [👶🤖🖥️ BabyAGI UI](https://github.com/miurla/babyagi-ui)                                    | BabyAGI UI is designed to make it easier to run and develop with babyagi in a web app, like a ChatGPT.                                                                                      |           |             |   Tool   |\n| [baichuan-7B](https://github.com/baichuan-inc/baichuan-7B)                                     | A large-scale 7B pretraining language model developed by Baichuan.                                                                                                                             |           |             |   Tool   |\n| [Baichuan-13B](https://github.com/baichuan-inc/Baichuan-13B)                                   | A 13B large language model developed by Baichuan Intelligent Technology.                                                                                                                       |          |              |   Tool   |\n| [Baichuan 2](https://github.com/baichuan-inc/Baichuan2)                                        | A series of large language models developed by Baichuan Intelligent Technology.                                                                                                                |           |             |   Tool   |\n| [Bisheng](https://github.com/dataelement/bisheng)                                              | Bisheng is an open LLM devops platform for next generation AI applications.                                                                                                                    |           |             |   Tool   |\n| [Character-LLM](https://github.com/choosewhatulike/trainable-agents)                           | A Trainable Agent for Role-Playing.                                                                                              |[arXiv](https://arxiv.org/abs/2310.10158)  |             |   Tool   |\n| [ChatDev](https://github.com/OpenBMB/ChatDev)                                                  | Communicative Agents for Software Development.                                                                                   |[arXiv](https://arxiv.org/abs/2307.07924)  |             |   Tool   |\n| [ChatGPT-API-unity](https://github.com/mochi-neko/ChatGPT-API-unity)                           | Binds ChatGPT chat completion API to pure C# on Unity.                                                                                                                                         |          |     Unity    |   Tool   |\n| [ChatGPTForUnity](https://github.com/sunsvip/ChatGPTForUnity)                                  | ChatGPT for unity.                                                                                                                                                                             |           |    Unity    |   Tool   |\n| [ChatRWKV](https://github.com/BlinkDL/ChatRWKV)                                                | ChatRWKV is like ChatGPT but powered by RWKV (100% RNN) language model, and open source.                                                                                                       |           |             |   Tool   |\n| [ChatYuan](https://github.com/clue-ai/ChatYuan)                                                | Large Language Model for Dialogue in Chinese and English.                                                                                                                                      |           |             |   Tool   |\n| [Chinese-LLaMA-Alpaca-3](https://github.com/ymcui/Chinese-LLaMA-Alpaca-3)                      | (Chinese Llama-3 LLMs) developed from Meta Llama 3.                                                                                                                                            |            |            |   Tool   |\n| [Chrome-GPT](https://github.com/richardyc/Chrome-GPT)                                          | An AutoGPT agent that controls Chrome on your desktop.                                                                                                                                         |           |             |   Tool   |\n| [CogVLM](https://www.modelscope.cn/models/ZhipuAI/CogVLM/summary)                              | CogVLM, a powerful open-source visual language foundation model.                                                                 |[arXiv](https://arxiv.org/abs/2311.03079)  |             |   Tool   |\n| [CoreNet](https://github.com/apple/corenet)                                                    | A library for training deep neural networks.                                                                                                                                                   |            |            |   Tool   |\n| [Cosmos](https://github.com/NVIDIA/Cosmos)                                                     | Cosmos is a world model development platform that consists of world foundation models, tokenizers and video processing pipeline to accelerate the development of Physical AI at Robotics \u0026 AV labs.      |            |             |   LLM   |\n| [DBRX](https://github.com/databricks/dbrx)                                                     | DBRX is a large language model trained by Databricks.                                                                                                                                          |          |              |   Tool   |\n| [DCLM](https://github.com/mlfoundations/dclm)                                                  | DataComp for Language Models.                                                                                                    |[arXiv](https://arxiv.org/abs/2406.11794)  |             |   Tool   |\n| [DeepSeek-R1](https://github.com/deepseek-ai/DeepSeek-R1)                                      | DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning.       |             |             |   LLM   |\n| [DeepSeek-V3](https://github.com/deepseek-ai/DeepSeek-V3)                                      | DeepSeek-V3 is a strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token.             |[arXiv](https://arxiv.org/abs/2412.19437)  |             |   LLM   |\n| [DemoGPT](https://github.com/melih-unsal/DemoGPT)                                              | Auto Gen-AI App Generator with the Power of Llama 2                                                                                                                                            |          |              |   Tool   |\n| [Design2Code](https://github.com/NoviScl/Design2Code)                                          | Automating Front-End Engineering                                                                                                                                                               |          |              |   Tool   |\n| [Devika](https://github.com/stitionai/devika)                                                  | Devika is an Agentic AI Software Engineer.                                                                                                                                                     |          |              |   Tool   |\n| [Devon](https://github.com/entropy-research/Devon)                                             | An open-source pair programmer.                                                                                                                                                                |          |              |   Tool   |\n| [Dora](https://www.dora.run/ai)                                                                | Generating powerful websites, one prompt at a time.                                                                                                                                            |           |             |   Tool   |\n| [Flowise](https://github.com/FlowiseAI/Flowise)                                                | Drag \u0026 drop UI to build your customized LLM flow using LangchainJS.                                                                                                                            |           |             |   Tool   |\n| [Gemini](https://deepmind.google/technologies/gemini)                                          | Gemini is built from the ground up for multimodality — reasoning seamlessly across text, images, video, audio, and code.                                                                      |          |              |   Tool   |\n| [Gemma](https://github.com/google/gemma_pytorch)                                               | Gemma is a family of lightweight, state-of-the art open models built from research and technology used to create Google Gemini models.                                                      |          |              |   Tool   |\n| [gemma.cpp](https://github.com/google/gemma.cpp)                                               | lightweight, standalone C++ inference engine for Google's Gemma models.                                                                                                                        |          |              |   Tool   |\n| [GLM-4](https://github.com/THUDM/GLM-4)                                                        | GLM-4-9B is the open-source version of the latest generation of pre-trained models in the GLM-4 series launched by Zhipu AI.                                                                   |          |              |   Tool   |\n| [GLM-4.5](https://github.com/zai-org/GLM-4.5)                                                  | GLM-4.5: An open-source large language model designed for intelligent agents by Z.ai.                                                                                                          |          |              |   LLM   |\n| [GPT4All](https://github.com/nomic-ai/gpt4all)                                                 | A chatbot trained on a massive collection of clean assistant data including code, stories and dialogue.                                                                                        |           |             |   Tool   |\n| [GPT-4o](https://openai.com/index/hello-gpt-4o/)                                               | GPT-4o (“o” for “omni”) is a step towards much more natural human-computer interaction—it accepts as input any combination of text, audio, image, and video and generates any combination of text, audio, and image outputs.                                                                                                                                                                |          |              |   Tool   |\n| [gpt-oss](https://github.com/openai/gpt-oss)                                                   | gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI.                                                                                                                    |          |              |   LLM   |\n| [GPTScript](https://github.com/gptscript-ai/gptscript)                                         | Develop LLM Apps in Natural Language.                                                                                                                                                          |          |              |   Tool   |\n| [Grok-1](https://x.ai/blog/grok-os)                                                            | The weights and architecture of our 314 billion parameter Mixture-of-Experts model, Grok-1.                                                                                                   |          |              |   Tool   |\n| [HuggingChat](https://huggingface.co/chat/)                                                    | Making the community's best AI chat models available to everyone.                                                                                                                              |          |              |   Tool   |\n| [Hugging Face API Unity Integration](https://github.com/huggingface/unity-api)                 | This Unity package provides an easy-to-use integration for the Hugging Face Inference API, allowing developers to access and use Hugging Face AI models within their Unity projects.       |          |     Unity     |   Tool   |\n| [Hunyuan-MT](https://github.com/Tencent-Hunyuan/Hunyuan-MT)                                    | The Hunyuan-MT comprises a translation model, Hunyuan-MT-7B, and an ensemble model, Hunyuan-MT-Chimera. The translation model is used to translate source text into the target language, while the ensemble model integrates multiple translation outputs to produce a higher-quality result.                                                                                      |          |              |   LLM   |\n| [ImageBind](https://github.com/facebookresearch/ImageBind)                                     | ImageBind One Embedding Space to Bind Them All.                                                                                       |[arXiv](https://arxiv.org/abs/2305.05665)  |        |   Tool   |\n| [Index-1.9B](https://github.com/bilibili/Index-1.9B)                                           | A SOTA lightweight multilingual LLM.                                                                                                                                                            |          |              |   Tool   |\n| [InteractML-Unity](https://github.com/Interactml/iml-unity)                                    | InteractML, an Interactive Machine Learning Visual Scripting framework for Unity3D.                                                                                                            |          |     Unity     |   Tool   |\n| [InteractML-Unreal Engine](https://github.com/Interactml/iml-ue4)                              | Bringing Machine Learning to Unreal Engine.                                                                                                                                                    |          | Unreal Engine |   Tool   |\n| [InternLM](https://github.com/InternLM/InternLM)                                               | InternLM has open-sourced a 7 billion parameter base model, a chat model tailored for practical scenarios and the training system.   |[arXiv](https://arxiv.org/abs/2403.17297)  |     |   Tool   |\n| [InternLM-XComposer](https://github.com/InternLM/InternLM-XComposer)                           | InternLM-XComposer2 is a groundbreaking vision-language large model (VLLM) excelling in free-form text-image composition and comprehension.  |[arXiv](https://arxiv.org/abs/2404.06512)  |     |   Tool   |\n| [Jan](https://github.com/janhq/jan)                                                            | Bring AI to your Desktop.                                                                                                                                                                      |          |              |   Tool   |\n| [Janus](https://github.com/deepseek-ai/Janus)                                                  | Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.                                                   |[arXiv](https://arxiv.org/abs/2410.13848)  |     |   LLM   |\n| [Kimi K2](https://github.com/moonshotai/Kimi-K2)                                               | Kimi K2 is a state-of-the-art mixture-of-experts (MoE) language model with 32 billion activated parameters and 1 trillion total parameters.                                                |          |              |   LLM   |\n| [Lamini](https://github.com/lamini-ai/lamini)                                                  | Lamini allows any engineering team to outperform general purpose LLMs through RLHF and fine- tuning on their own data.                                                                      |          |              |   Tool   |\n| [LaMini-LM](https://github.com/mbzuai-nlp/LaMini-LM)                                           | LaMini-LM is a collection of small-sized, efficient language models distilled from ChatGPT and trained on a large-scale dataset of 2.58M instructions.                                  |          |              |   Tool   |\n| [LangChain](https://github.com/hwchase17/langchain)                                            | LangChain is a framework for developing applications powered by language models.                                                                                                               |          |              |   Tool   |\n| [LangFlow](https://github.com/logspace-ai/langflow)                                            | ⛓️ LangFlow is a UI for LangChain, designed with react-flow to provide an effortless way to experiment and prototype flows.                                                                   |          |              |   Tool   |\n| [LaVague](https://github.com/lavague-ai/LaVague)                                               | Automate automation with Large Action Model framework.                                                                                                                                         |          |              |   Tool   |\n| [Lemur](https://github.com/OpenLemur/Lemur)                                                    | Open Foundation Models for Language Agents.                                                                                                                                                    |          |              |   Tool   |\n| [Lepton AI](https://github.com/leptonai/leptonai)                                              | A Pythonic framework to simplify AI service building.                                                                                                                                          |          |              |   Tool   |\n| [Lit-LLaMA](https://github.com/Lightning-AI/lit-llama)                                         | Implementation of the LLaMA language model based on nanoGPT. Supports flash attention, Int8 and GPTQ 4bit quantization, LoRA and LLaMA-Adapter fine-tuning, pre-training.                   |          |              |   Tool   |\n| [llama2-webui](https://github.com/liltom-eth/llama2-webui)                                     | Run Llama 2 locally with gradio UI on GPU or CPU from anywhere (Linux/Windows/Mac).                                                                                                            |          |              |   Tool   |\n| [Llama 3](https://github.com/meta-llama/llama3)                                                | The official Meta Llama 3 GitHub site.                                                                                                                                                         |          |              |   Tool   |\n| [Llama 3.1](https://github.com/meta-llama/llama-models)                                        | Llama is an accessible, open large language model (LLM) designed for developers, researchers, and businesses to build, experiment, and responsibly scale their generative AI ideas.                                                                                                                                                         |          |              |   Tool   |\n| [LLaSM](https://github.com/LinkSoul-AI/LLaSM)                                                  | Large Language and Speech Model.                                                                                                                                                               |          |              |   Tool   |\n| [LLM Answer Engine](https://github.com/developersdigest/llm-answer-engine)                     | Build a Perplexity-Inspired Answer Engine Using Next.js, Groq, Mixtral, Langchain, OpenAI, Brave \u0026 Serper.                                                                              |           |             |   Tool   |\n| [llm.c](https://github.com/karpathy/llm.c)                                                     | LLM training in simple, raw C/CUDA.                                                                                                                                                            |          |              |   Tool   |\n| [LLMUnity](https://github.com/undreamai/LLMUnity)                                              | Create characters in Unity with LLMs!                                                                                                                                                          |          |     Unity    |   Tool   |\n| [LLocalSearch](https://github.com/nilsherzig/LLocalSearch)                                     | LLocalSearch is a completely locally running search engine using LLM Agents.                                                                                                                   |          |              |   Tool   |\n| [LogicGamesSolver](https://github.com/fabridigua/LogicGamesSolver)                             | A Python tool to solve logic games with AI, Deep Learning and Computer Vision.                                                                                                                 |          |              |   Tool   |\n| [LongCat-Flash](https://github.com/meituan-longcat/LongCat-Flash-Chat)                         | LongCat-Flash is a powerful and efficient language model with 560 billion total parameters, featuring an innovative Mixture-of-Experts (MoE) architecture. The model incorporates a dynamic computation mechanism that activates 18.6B∼31.3B parameters (averaging∼27B) based on contextual demands, optimizing both computational efficiency and performance.           |          |              |   LLM   |\n| [LongWriter](https://github.com/THUDM/LongWriter)                                              | LongWriter: Unleashing 10,000+ Word Generation From Long Context LLMs.                                                          |[arXiv](https://arxiv.org/abs/2408.07055)  |              |   Tool   |\n| [Large World Model (LWM)](https://github.com/LargeWorldModel/LWM)                              | Large World Model (LWM) is a general-purpose large-context multimodal autoregressive model.                                |[arXiv](https://arxiv.org/abs/2402.08268)  |              |   Tool   |\n| [Lumina-T2X](https://github.com/Alpha-VLLM/Lumina-T2X)                                         | Lumina-T2X is a unified framework for Text to Any Modality Generation.                                                          |[arXiv](https://arxiv.org/abs/2405.05945)  |              |   Tool   |\n| [MetaGPT](https://github.com/geekan/MetaGPT)                                                   | The Multi-Agent Framework                                                                                                                                                                      |          |              |   Tool   |\n| [MiniCPM-2B](https://github.com/OpenBMB/MiniCPM)                                               | An end-side LLM outperforms Llama2-13B.                                                                                                                                                        |          |              |   Tool   |\n| [MiniGPT-4](https://github.com/Vision-CAIR/MiniGPT-4)                                          | Enhancing Vision-language Understanding with Advanced Large Language Models.                                                    |[arXiv](https://arxiv.org/abs/2304.10592)  |              |   Tool   |\n| [MiniGPT-5](https://github.com/eric-ai-lab/MiniGPT-5)                                          | Interleaved Vision-and-Language Generation via Generative Vokens.                                                               |[arXiv](https://arxiv.org/abs/2310.02239)  |              |   Tool   |\n| [MiniMax-01](https://github.com/MiniMax-AI/MiniMax-01)                                         | MiniMax-01: Scaling Foundation Models with Lightning Attention.                                                                 |[arXiv](https://arxiv.org/abs/2501.08313)  |              |   LLM   |\n| [Mixtral 8x7B](https://mistral.ai/news/mixtral-of-experts/)                                    | A high quality Sparse Mixture-of-Experts.                                                                                       |[arXiv](https://arxiv.org/abs/2401.04088)  |              |   Tool   |\n| [Mistral 7B](https://mistral.ai/news/announcing-mistral-7b/)                                   | The best 7B model to date, Apache 2.0.                                                                                                                                                         |          |              |   Tool   |\n| [Mistral Large](https://mistral.ai/news/mistral-large/)                                        | Mistral Large is a new cutting-edge text generation model. It reaches top-tier reasoning capabilities.                                                                                         |          |              |   Tool   |\n| [MLC LLM](https://github.com/mlc-ai/mlc-llm)                                                   | Enable everyone to develop, optimize and deploy AI models natively on everyone's devices.                                                                                                      |          |              |   Tool   |\n| [MobiLlama](https://github.com/mbzuai-oryx/MobiLlama)                                          | Towards Accurate and Lightweight Fully Transparent GPT.                                                                         |[arXiv](https://arxiv.org/abs/2402.16840)  |              |   Tool   |\n| [MoE-LLaVA](https://github.com/PKU-YuanGroup/MoE-LLaVA)                                        | Mixture of Experts for Large Vision-Language Models.                                                                                |[arXiv](https://arxiv.org/abs/2401.15947)  |              |   Tool   |\n| [Moshi](https://www.moshi.chat/?queue_id=talktomoshi)                                          | Moshi is an experimental conversational AI.                                                                                                                                                    |          |              |   Tool   |\n| [Moshi](https://github.com/kyutai-labs/moshi)                                                  | Moshi: a speech-text foundation model for real time dialogue.                                                                                                                                                    |          |              |   Tool   |\n| [MOSS](https://github.com/OpenLMLab/MOSS)                                                      | An open-source tool-augmented conversational language model from Fudan University.                                                                                                             |          |              |   Tool   |\n| [mPLUG-Owl🦉](https://github.com/X-PLUG/mPLUG-Owl)                                            | Modularization Empowers Large Language Models with Multimodality.                                                               |[arXiv](https://arxiv.org/abs/2304.14178)  |              |   Tool   |\n| [Nemotron-4](https://arxiv.org/abs/2402.16819)                                                 | A 15-billion-parameter large multilingual language model trained on 8 trillion text tokens.                               |[arXiv](https://arxiv.org/abs/2402.16819)  |              |   Tool   |\n| [NExT-GPT](https://github.com/NExT-GPT/NExT-GPT)                                               | Any-to-Any Multimodal Large Language Model.                                                                                                                                                    |          |              |   Tool   |\n| [OLMo](https://github.com/allenai/OLMo)                                                        | Open Language Model                                                                                                              |[arXiv](https://arxiv.org/abs/2402.00838)  |             |   Tool   |\n| [OmniLMM](https://github.com/OpenBMB/OmniLMM)                                                  | Large multi-modal models for strong performance and efficient deployment.                                                                                                                      |          |              |   Tool   |\n| [OneLLM](https://github.com/csuhan/OneLLM)                                                     | One Framework to Align All Modalities with Language.                                                                            |[arXiv](https://arxiv.org/abs/2312.03700)  |              |   Tool   |\n| [Open-Assistant](https://github.com/LAION-AI/Open-Assistant)                                   | OpenAssistant is a chat-based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so.                                        |          |              |   Tool   |\n| [Open Deep Research](https://github.com/dzhng/deep-research)                                   | An AI-powered research assistant that performs iterative, deep research on any topic by combining search engines, web scraping, and large language models.                                   |          |              |   LLM   |\n| [OpenDevin](https://github.com/OpenDevin/OpenDevin)                                            | An autonomous AI software engineer.                                                                                                                                                            |          |              |   Tool   |\n| [Orion-14B](https://github.com/OrionStarAI/Orion)                                              | Orion-14B is a family of models includes a 14B foundation LLM, and a series of models.                                          |[arXiv](https://arxiv.org/abs/2401.12246)  |              |   Tool   |\n| [Panda](https://github.com/dandelionsllm/pandallm)                                             | Overseas Chinese open source large language model, based on Llama-7B, -13B, -33B, -65B for continuous pre-training in the Chinese field.                                                    |          |              |   Tool   |\n| [Perplexica](https://github.com/ItzCrazyKns/Perplexica)                                        | An AI-powered search engine.                                                                                                                                                                   |           |             |   Tool   |\n| [Pi](https://heypi.com/talk)                                                                   | AI chatbot designed for personal assistance and emotional support.                                                                                                                             |          |              |   Tool   |\n| [Qwen1.5](https://github.com/QwenLM/Qwen1.5)                                                   | Qwen1.5 is the improved version of Qwen.                                                                                                                                                       |           |             |   Tool   |\n| [Qwen2](https://github.com/QwenLM/Qwen2)                                                       | Qwen2 is the large language model series developed by Qwen team, Alibaba Cloud.                                                                                                                |           |             |   LLM   |\n| [Qwen2.5-Coder](https://github.com/QwenLM/Qwen2.5-Coder)                                       | Qwen2.5-Coder is the code version of Qwen2.5, the large language model series developed by Qwen team, Alibaba Cloud.                         |[arXiv](https://arxiv.org/abs/2409.12186)  |             |   LLM   |\n| [Qwen-7B](https://github.com/QwenLM/Qwen-7B)                                                   | The official repo of Qwen-7B (通义千问-7B) chat \u0026 pretrained large language model proposed by Alibaba Cloud.                                                                                    |          |              |   LLM   |\n| [Qwen3](https://github.com/QwenLM/Qwen3)                                                       | Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.                                                                |[arXiv](https://arxiv.org/abs/2505.09388)  |              |   LLM   |\n| [RepoAgent](https://github.com/OpenBMB/RepoAgent)                                              | RepoAgent is an Open-Source project driven by Large Language Models(LLMs) that aims to provide an intelligent way to document projects.     |[arXiv](https://arxiv.org/abs/2402.16667)  |              |   Tool   |\n| [s1](https://github.com/simplescaling/s1)                                                      | s1: Simple test-time scaling.                                                                                                                  |[arXiv](https://arxiv.org/abs/2501.19393)  |              |   LLM   |\n| [Sanity AI Engine](https://github.com/tosos/SanityEngine)                                      | Sanity AI Engine for the Unity Game Development Tool.                                                                                                                                          |          |     Unity     |   Tool   |\n| [SearchGPT](https://github.com/tobiasbueschel/search-gpt)                                      | 🌳 Connecting ChatGPT with the Internet                                                                                                                                                       |          |              |   Tool   |\n| [Seed-OSS](https://github.com/ByteDance-Seed/seed-oss)                                         | Seed-OSS is a series of open-source large language models developed by ByteDance's Seed Team, designed for powerful long-context, reasoning, agent and general capabilities, and versatile developer-friendly features.                                                                                                                                                                   |          |              |   LLM   |\n| [ShareGPT4V](https://sharegpt4v.github.io/)                                                    | Improving Large Multi-Modal Models with Better Captions.                                                                                                                                       |          |              |   Tool   |\n| [SkyThought](https://github.com/NovaSky-AI/SkyThought)                                         | Sky-T1: Train your own O1 preview model within $450.                                                                                                                                           |          |              |   LLM   |\n| [Skywork](https://github.com/SkyworkAI/Skywork)                                                | Skywork series models are pre-trained on 3.2TB of high-quality multilingual (mainly Chinese and English) and code data.                                                                  |          |              |   Tool   |\n| [StableLM](https://github.com/Stability-AI/StableLM)                                           | Stability AI Language Models.                                                                                                   |[arXiv](https://arxiv.org/abs/2402.17834)  |              |   Tool   |\n| [Stanford Alpaca](https://github.com/tatsu-lab/stanford_alpaca)                                | An Instruction-following LLaMA Model.                                                                                                                                                          |          |              |   LLM   |\n| [Text generation web UI](https://github.com/oobabooga/text-generation-webui)                   | A gradio web UI for running Large Language Models like LLaMA, llama.cpp, GPT-J, OPT, and GALACTICA.                                                                                           |          |              |   Tool   |\n| [TinyChatEngine](https://github.com/mit-han-lab/TinyChatEngine)                                | On-Device LLM Inference Library.                                                                                                                                                               |          |              |   Tool   |\n| [ToolBench](https://github.com/openbmb/toolbench)                                              | An open platform for training, serving, and evaluating large language model for tool learning.                                                                                            |           |             |   Tool   |\n| [Unity ChatGPT](https://github.com/dilmerv/UnityChatGPT)                                       | Unity ChatGPT Experiments.                                                                                                                                                                     |          |     Unity     |   Tool   |\n| [Unity OpenAI-API Integration](https://github.com/himanshuskyrockets/Unity_OpenAI)             | Integrate openai GPT-3 language model and ChatGPT API into a Unity project.                                                                                                                    |          |     Unity     |   Tool   |\n| [Unreal Engine 5 Llama LoRA](https://github.com/bublint/ue5-llama-lora)                        | A proof-of-concept project that showcases the potential for using small, locally trainable LLMs to create next-generation documentation tools.                                        |          | Unreal Engine |   Tool   |\n| [UnrealGPT](https://github.com/TREE-Ind/UnrealGPT)                                             | A collection of Unreal Engine 5 Editor Utility widgets powered by GPT3/4.                                                                                                                      |          | Unreal Engine |   Tool   |\n| [Video-LLaVA](https://github.com/PKU-YuanGroup/Video-LLaVA)                                    | Learning United Visual Representation by Alignment Before Projection.                                                           |[arXiv](https://arxiv.org/abs/2311.10122)  |              |   Tool   |\n| [WebGPT](https://github.com/0hq/WebGPT)                                                        | Run GPT model on the browser with WebGPU.                                                                                                                                                      |          |              |   Tool   |\n| [Web3-GPT](https://github.com/Markeljan/Web3GPT)                                               | Deploy smart contracts with AI                                                                                                                                                                 |          |              |   Tool   |\n| [WordGPT](https://github.com/filippofinke/WordGPT)                                             | 🤖 Bring the power of ChatGPT to Microsoft Word                                                                                                                                               |          |              |   Tool   |\n| [XAgent](https://github.com/OpenBMB/XAgent)                                                    | An Autonomous LLM Agent for Complex Task Solving.                                                                                                                                              |          |              |   Tool   |\n| [Yi](https://github.com/01-ai/Yi)                                                              | A series of large language models trained from scratch by developers.                                                                                                                          |          |              |   Tool   |\n| [01 Project](https://github.com/OpenInterpreter/01)                                            | The open-source language model computer.                                                                                                                                                       |          |              |   Tool   | \n| [SimpleOllamaUnity](https://github.com/HardCodeDev777/SimpleOllamaUnity)                       | Ollama integration for Unity Engine (works in runtime and editor)                                                                                                                              |          |     Unity    |   Tool   |\n| [AI-Writer](https://github.com/BlinkDL/AI-Writer)                                              | AI writes novels, generates fantasy and romance web articles, etc. Chinese pre-trained generative model.                                                                                    |               |              |  Writer  |\n| [Notebook.ai](https://github.com/indentlabs/notebook)                                          | Notebook.ai is a set of tools for writers, game designers, and roleplayers to create magnificent universes – and everything within them.                                                  |               |              |  Writer  |\n| [Novel](https://github.com/steven-tey/novel)                                                   | Notion-style WYSIWYG editor with AI-powered autocompletions.                                                                                                                                   |               |              |  Writer  |\n| [NovelAI](https://novelai.net/)                                                                | Driven by AI, painlessly construct unique stories, thrilling tales, seductive romances, or just fool around.                                                                                 |               |              |  Writer  |\n| [Unity-MCP](https://github.com/IvanMurzak/Unity-MCP) | Open-source MCP server connecting AI agents to the Unity Editor and runtime, with 100+ built-in tools. |  | Unity | Tool |\n| [Godot-MCP](https://github.com/IvanMurzak/Godot-MCP) | Open-source MCP server connecting AI agents to the Godot Editor and runtime (Godot 4.x, C#). |  | Godot | Tool |\n| [Unreal-MCP](https://github.com/IvanMurzak/Unreal-MCP) | Open-source MCP server connecting AI agents to Unreal Engine 5.7, editor and runtime (C++ plugin + .NET sidecar). |  | Unreal Engine | Tool |\n| [GameDev-MCP-Server](https://github.com/IvanMurzak/GameDev-MCP-Server) | Open-source, engine-agnostic MCP server shared by Unity-MCP, Godot-MCP, and Unreal-MCP. |  | Unity/Godot/Unreal Engine | Tool |\n| [MCP-Plugin-dotnet](https://github.com/IvanMurzak/MCP-Plugin-dotnet) | Open-source .NET library/SDK that turns any .NET application into an MCP server. |  |  | Tool |\n| [ReflectorNet](https://github.com/IvanMurzak/ReflectorNet) | Open-source .NET reflection toolkit for AI-driven scenarios. |  |  | Tool |\n\n\n\u003cp style=\"text-align: right;\"\u003e\u003ca href=\"#table-of-contents\"\u003e^ Back to Contents ^\u003c/a\u003e\u003c/p\u003e\n\n\n## \u003cspan id=\"visual\"\u003eVLM (Visual)\u003c/span\u003e\n\n| Source                                                                                      | Description                                                                                                                                                                                    |   Paper   |  Game Engine  |   Type   |\n| :------------------------------------------------------------------------------------------ | :--------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | :-----------: | :-----------: | :-------: |\n| [Cambrian-1](https://github.com/cambrian-mllm/cambrian)                                     | Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs.                                                            |[arXiv](https://arxiv.org/abs/2406.16860)  |              |   Multimodal LLMs  |\n| [CogVLM2](https://github.com/THUDM/CogVLM2)                                                 | GPT4V-level open-source multi-modal model based on Llama3-8B.                                                                                       |                           |              |   Visual  |\n| [CoTracker](https://co-tracker.github.io/)                                                  | It is Better to Track Together.                                                                                                      |[arXiv](https://arxiv.org/abs/2307.07635)  |               | Visual |\n| [dots.vlm1](https://github.com/rednote-hilab/dots.vlm1)                                     | dots.vlm1 is the first vision-language model in the dots model family. Built upon a 1.2 billion-parameter vision encoder and the DeepSeek V3 large language model (LLM), dots.vlm1 demonstrates strong multimodal understanding and reasoning capabilities.                                                                                 |                           |              |   VLM  |\n| [EVF-SAM](https://github.com/hustvl/EVF-SAM)                                                | EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model.                                                      |[arXiv](https://arxiv.org/abs/2406.20076)  |               | Visual |\n| [FaceHi](https://m.facehi.ai/)                                                              | It is Better to Track Together.                                                                                                                       |                                           |               | Visual |\n| [GLM-V](https://github.com/zai-org/GLM-V)                                                   | GLM-4.1V-Thinking and GLM-4.5V: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning.                 |[arXiv](https://arxiv.org/abs/2507.01006)  |               | VLM |\n| [InternLM-XComposer2](https://github.com/InternLM/InternLM-XComposer)                       | InternLM-XComposer2 is a groundbreaking vision-language large model (VLLM) excelling in free-form text-image composition and comprehension.           |[arXiv](https://arxiv.org/abs/2404.06512)  |               | Visual |\n| [Kangaroo](https://github.com/KangarooGroup/Kangaroo)                                       | Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input.                                                                        |                                           |               | Visual |\n| [Kwai Keye-VL](https://github.com/Kwai-Keye/Keye)                                           | Kwai Keye-VL is a cutting-edge multimodal large language model meticulously crafted by the Kwai Keye Team at Kuaishou.              |[arXiv](https://arxiv.org/abs/2509.01563)  |               | VLM |\n| [LGVI](https://jianzongwu.github.io/projects/rovi/)                                         | Towards Language-Driven Video Inpainting via Multimodal Large Language Models.                                                                         |                                           |               | Visual |\n| [LLaVA++](https://github.com/mbzuai-oryx/LLaVA-pp)                                          | Extending Visual Capabilities with LLaMA-3 and Phi-3.                                                                                                     |                                     |              |   Visual  |\n| [LLaVA-OneVision](https://github.com/LLaVA-VL/LLaVA-NeXT)                                   | LLaVA-OneVision: Easy Visual Task Transfer.                                                                                           |[arXiv](https://arxiv.org/abs/2408.03326)  |              |   Visual  |\n| [LongVA](https://github.com/EvolvingLMMs-Lab/LongVA)                                        | Long Context Transfer from Language to Vision.                                                                                        |[arXiv](https://arxiv.org/abs/2406.16852)  |              |   Visual  |\n| [Lumina-DiMOO](https://github.com/Alpha-VLLM/Lumina-DiMOO)                                  | Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding.                                                |                                     |              |   VLM  |\n| [MaskViT](https://maskedvit.github.io/)                                                     | Masked Visual Pre-Training for Video Prediction.                                                                                      |[arXiv](https://arxiv.org/abs/2206.11894)  |              | Visual |\n| [MiniCPM-Llama3-V 2.5](https://github.com/OpenBMB/MiniCPM-V)                                | A GPT-4V Level MLLM on Your Phone.                                                                                                                        |                                      |              |   Visual  |\n| [MiniCPM-V 4.0](https://github.com/OpenBMB/MiniCPM-o)                                       | MiniCPM-V 4.0: A GPT-4V Level MLLM for Single Image, Multi Image and Video on Your Phone.                                                                 |                                      |              |   Visual  |\n| [MoE-LLaVA](https://github.com/PKU-YuanGroup/MoE-LLaVA)                                     | Mixture of Experts for Large Vision-Language Models.                                                                                  |[arXiv](https://arxiv.org/abs/2401.15947)  |              |   Visual  |\n| [MotionLLM](https://github.com/IDEA-Research/MotionLLM)                                     | Understanding Human Behaviors from Human Motions and Videos.                                                                          |[arXiv](https://arxiv.org/abs/2405.20340)  |              |   Visual  |\n| [PLLaVA](https://github.com/magic-research/PLLaVA)                                          | Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning.                                                      |[arXiv](https://arxiv.org/abs/2404.16994)  |              |   Visual  |\n| [POINTS-Reader](https://github.com/Tencent/POINTS-Reader)                                   | POINTS-Reader: Distillation-Free Adaptation of Vision-Language Models for Document Conversion.                              |[arXiv](https://arxiv.org/abs/2509.01215)  |              |   Visual  |\n| [Qwen-VL](https://github.com/QwenLM/Qwen-VL)                                                | A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.                                          |[arXiv](https://arxiv.org/abs/2308.12966)  |              |   Visual  |\n| [Sapiens](https://github.com/facebookresearch/sapiens)                                      | Sapiens: Foundation for Human Vision Models.                                                                                          |[arXiv](https://arxiv.org/abs/2408.12569)  |              |   Visual  |\n| [ShareGPT4V](https://github.com/ShareGPT4Omni/ShareGPT4V)                                   | Improving Large Multi-modal Models with Better Captions.                                                                              |[arXiv](https://arxiv.org/abs/2311.12793)  |              |   Visual  |\n| [SOLO](https://github.com/Yangyi-Chen/SOLO)                                                 | SOLO: A Single Transformer for Scalable Vision-Language Modeling.                                                                     |[arXiv](https://arxiv.org/abs/2407.06438)  |              |   Visual  |\n| [VideoAgent](https://github.com/YueFan1014/VideoAgent)                                      | VideoAgent: A Memory-augmented Multimodal Agent for Video Understanding.                                                              |[arXiv](https://arxiv.org/abs/2403.11481)  |              |   Agent  |\n| [Video-CCAM](https://github.com/QQ-MM/Video-CCAM)                                           | Video-CCAM: Advancing Video-Language Understanding with Causal Cross-Attention Masks.                                                                                          |  |              |   Visual  |\n| [Video-LLaVA](https://github.com/PKU-YuanGroup/Video-LLaVA)                                 | Learning United Visual Representation by Alignment Before Projection.                                                                 |[arXiv](https://arxiv.org/abs/2311.10122)  |              |   Visual  |\n| [VideoLLaMA 2](https://github.com/DAMO-NLP-SG/VideoLLaMA2)                                  | Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.                                                            |[arXiv](https://arxiv.org/abs/2406.07476)  |              |   Visual  |\n| [VideoLLaMA 3](https://github.com/DAMO-NLP-SG/VideoLLaMA3)                                  | VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding.                                                |[arXiv](https://arxiv.org/abs/2501.13106)  |              |   Visual  |\n| [Video-MME](https://github.com/BradyFU/Video-MME)                                           | The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.                                              |[arXiv](https://arxiv.org/abs/2405.21075)  |              |   Visual  |\n| [Vitron](https://github.com/SkyworkAI/Vitron)                                               | A Unified Pixel-level Vision LLM for Understanding, Generating, Segmenting, Editing.                                                                      |                                      |              |   Visual  |\n| [VILA](https://github.com/NVlabs/VILA)                                                      | VILA: On Pre-training for Visual Language Models.                                                                                     |[arXiv](https://arxiv.org/abs/2312.07533)  |              |   Visual  |\n\n\u003cp style=\"text-align: right;\"\u003e\u003ca href=\"#table-of-contents\"\u003e^ Back to Contents ^\u003c/a\u003e\u003c/p\u003e\n\n\n## \u003cspan id=\"game\"\u003eGame (World Model \u0026 Agent)\u003c/span\u003e\n\n| Source                                                                                      | Description                                                                                                                                                                                    |   Paper   |  Game Engine  |   Type   |\n| :------------------------------------------------------------------------------------------ | :--------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | :-----------: | :-----------: | :-------: |\n| [AgentBench](https://github.com/thudm/agentbench)                                              | A Comprehensive Benchmark to Evaluate LLMs as Agents.                                                                                 |[arXiv](https://arxiv.org/abs/2308.03688)  |        |   Agent  |\n| [Agent Group Chat](https://github.com/MikeGu721/AgentGroup)                                    | An Interactive Group Chat Simulacra For Better Eliciting Collective Emergent Behavior.                                                |[arXiv](https://arxiv.org/abs/2403.13433)  |        |   Agent  |\n| [Agent K](https://github.com/mikekelly/AgentK)                                                 | An autoagentic AGI that is self-evolving and modular.                                                                                                                            |         |              |   Agent  |\n| [Agent Laboratory](https://github.com/SamuelSchmidgall/AgentLaboratory)                        | Agent Laboratory: Using LLM Agents as Research Assistants.                                                                            |[arXiv](https://arxiv.org/abs/2501.04227)  |        |   Agent  |\n| [AgentScope](https://github.com/modelscope/agentscope)                                         | Start building LLM-empowered multi-agent applications in an easier way.                                                               |[arXiv](https://arxiv.org/abs/2402.14034)  |              |   Agent  |\n| [AgentSims](https://github.com/py499372727/AgentSims/)                                         | An Open-Source Sandbox for Large Language Model Evaluation.                                                                                                                            |         |              |   Agent  |\n| [AI Town](https://github.com/a16z-infra/ai-town)                                               | AI Town is a virtual town where AI characters live, chat and socialize.                                                                                                                |         |              |   Agent  |\n| [anime.gf](https://github.com/cyanff/anime.gf)                                                 | Local \u0026 Open Source Alternative to CharacterAI.                                                                                                                                         |        |              |   Game   |\n| [Astrocade](https://www.astrocade.com/)                                                        | Create games with AI                                                                                                                                                                    |        |              |   Game   |\n| [Atomic Agents](https://github.com/KennyVaneetvelde/atomic_agents)                             | The Atomic Agents framework is designed to be modular, extensible, and easy to use.                                                                                                     |        |              |   Agent  |\n| [AutoAgents](https://github.com/Link-AGI/AutoAgents)                                           | A Framework for Automatic Agent Generation.                                                                                                                                             |        |              |   Agent  |\n| [AutoGen](https://github.com/microsoft/autogen)                                                | Enable Next-Gen Large Language Model Applications.                                                                              |[arXiv](https://arxiv.org/abs/2308.08155)  |              |   Agent  |\n| [AWorld](https://github.com/inclusionAI/AWorld)                                             | AWorld: The Agent Runtime for Self-Improvement.                                                                                                                                            |        |              |   Agent  |\n| [behaviac](https://github.com/Tencent/behaviac)                                                | Behaviac is a framework of the game AI development.                                                                                                                               |              |              | Framework |\n| [Biomes](https://github.com/ill-inc/biomes-game)                                               | Biomes is an open source sandbox MMORPG built for the web using web technologies such as Next.js, Typescript, React and WebAssembly.                                                    |       |              |   Game   |\n| [Buffer of Thoughts](https://github.com/YangLing0818/buffer-of-thought-llm)                    | Thought-Augmented Reasoning with Large Language Models.                                                                         |[arXiv](https://arxiv.org/abs/2406.04271)  |              |   Agent  |\n| [Byzer-Agent](https://github.com/allwefantasy/byzer-agent)                                     | Easy, fast, and distributed agent framework for everyone.                                                                                                                               |        |              |   Agent  |\n| [Cat Town](https://github.com/ykhli/cat-town)                                                  | A C(h)atGPT-powered simulation with cats.                                                                                                                                               |        |              |   Agent  |\n| [Cat Town](https://github.com/ykhli/cat-town)                                                  | A C(h)atGPT-powered simulation with cats.                                                                                                                                               |        |              |   Agent  |\n| [CharacterGLM](https://github.com/thu-coai/CharacterGLM-6B)                                    | Customizing Chinese Conversational AI Characters with Large Language Models.                                                          |[arXiv](https://arxiv.org/abs/2311.16832)  |              |   Agent  |\n| [ChatDev](https://github.com/OpenBMB/ChatDev)                                                  | Communicative Agents for Software Development.                                                                                        |[arXiv](https://arxiv.org/abs/2405.04219)  |              |   Agent  |\n| [CogAgent](https://modelscope.cn/models/ZhipuAI/cogagent-chat/summary)                         | CogAgent is an open-source visual language model improved based on CogVLM.                                                            |[arXiv](https://arxiv.org/abs/2312.08914)  |              |   Agent  |\n| [ComoRAG](https://github.com/EternityJune25/ComoRAG)                                           | ComoRAG: A Cognitive-Inspired Memory-Organized RAG for Stateful Long Narrative Reasoning.                                             |[arXiv](https://arxiv.org/abs/2508.10419)  |              |   Agent  |\n| [Cradle](https://github.com/BAAI-Agents/Cradle)                                                | Towards General Computer Control.                                                                                                                                                         |      |              |   Agent  |\n| [crewAI](https://github.com/joaomdmoura/crewAI)                                                | Framework for orchestrating role-playing, autonomous AI agents.                                                                                                                          |       |              |   Agent  |\n| [Datarus Jupyter Agent](https://github.com/DatarusAI/Datarus-JupyterAgent)                     | The Datarus Jupyter Agent is a powerful multi-step reasoning system that executes complex analytical workflows with step-by-step reasoning, automatic error recovery, and comprehensive result synthesis.                                                                                                                                                                             |       |              |   Agent  |\n| [Dify](https://github.com/langgenius/dify)                                                     | Dify is an open-source LLM app building platform.                                                                                                                                        |       |              |   Agent  |\n| [Digital Life Project](https://digital-life-project.com/)                                      | Autonomous 3D Characters with Social Intelligence.                                                                                    |[arXiv](https://arxiv.org/abs/2312.04547)  |              |   Agent  |\n| [everything-ai](https://github.com/AstraBert/everything-ai)                                    | Your fully proficient, AI-powered and local chatbot assistant🤖.                                                                                                                        |       |              |   Agent  |\n| [fabric](https://github.com/danielmiessler/fabric)                                             | fabric is an open-source framework for augmenting humans using AI.                                                                                                                       |       |              |   Agent  |\n| [FastGPT](https://github.com/labring/FastGPT)                                                  | FastGPT is a knowledge-based platform built on the LLM.                                                                                                                                  |       |              |   Agent  |\n| [fastRAG](https://github.com/IntelLabs/fastRAG)                                                | Efficient Retrieval Augmentation and Generation Framework.                                                                                                                               |       |              |   Agent  |\n| [GameAISDK](https://github.com/Tencent/GameAISDK)                                              | Image-based game AI automation framework.                                                                                                                                         |              |              | Framework |\n| [GameNGen](https://gamengen.github.io/)                                                        | Diffusion Models Are Real-Time Game Engines.                                                                                          |[arXiv](https://arxiv.org/abs/2408.14837)  |              |   Game  |\n| [GameGen-O](https://github.com/GameGen-O/GameGen-O)                                            | GameGen-O: Open-world Video Game Generation.                                                                                                                                           |         |              |   Game   |\n| [GenAgent](https://github.com/xxyQwQ/GenAgent)                     | GenAgent: Build Collaborative AI Systems with Automated Workflow Generation - Case Studies on ComfyUI.                                                                                              |[arXiv](https://arxiv.org/abs/2409.01392)  |              |   Agent  |\n| [Generative Agents](https://github.com/joonspk-research/generative_agents)                     | Interactive Simulacra of Human Behavior.                                                                                              |[arXiv](https://arxiv.org/abs/2304.03442)  |              |   Agent  |\n| [Genesis](https://github.com/Genesis-Embodied-AI/Genesis)                                      | Genesis: A Generative and Universal Physics Engine for Robotics and Beyond.                                                                                                            |         |              |   Game   |\n| [Genie](https://sites.google.com/view/genie-2024/home)                                         | Generative Interactive Environments.                                                                                                                                                   |         |              |   Game   |\n| [Genie 3](https://deepmind.google/discover/blog/genie-3-a-new-frontier-for-world-models/)      | Genie 3: A new frontier for world models. Genie 3 is a general purpose world model that can generate an unprecedented diversity of interactive environments.                         |         |              |   Game   |\n| [gigax](https://github.com/GigaxGames/gigax)                                                   | Runtime, LLM-powered NPCs.                                                                                                                                                               |       |              |   Game   |\n| [HippoRAG](https://github.com/OSU-NLP-Group/HippoRAG)                                       | Neurobiologically Inspired Long-Term Memory for Large Language Models.                                                                   |[arXiv](https://arxiv.org/abs/2405.14831)  |              |   Agent   |\n| [Hunyuan-GameCraft](https://github.com/Tencent-Hunyuan/Hunyuan-GameCraft-1.0)               | Hunyuan-GameCraft: High-dynamic Interactive Game Video Generation with Hybrid History Condition.                                  |[arXiv](https://arxiv.org/abs/2506.17201)  |              |   Game   |\n| [HunyuanWorld 1.0](https://github.com/Tencent-Hunyuan/HunyuanWorld-1.0)                     | HunyuanWorld 1.0: Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels.                                     |[arXiv](https://arxiv.org/abs/2507.21809)  |              |   Game   |\n| [HunyuanWorld-Voyager](https://github.com/Tencent-Hunyuan/HunyuanWorld-Voyager)             | HunyuanWorld-Voyager is a novel video diffusion framework that generates world-consistent 3D point-cloud sequences from a single image with user-defined camera path. Voyager can generate 3D-consistent scene videos for world exploration following custom camera trajectories.                                                                                                          |       |              |   Game   |\n| [HY-World 1.5](https://github.com/Tencent-Hunyuan/HY-WorldPlay)                                | HY-World 1.5: A Systematic Framework for Interactive World Modeling with Real-Time Latency and Geometric Consistency.                                                               |       |              |   Game   |\n| [Interactive LLM Powered NPCs](https://github.com/AkshitIreddy/Interactive-LLM-Powered-NPCs)   | Interactive LLM Powered NPCs, is an open-source project that completely transforms your interaction with non-player characters (NPCs) in any game!                                       |        |              |   Game   |\n| [IoA](https://github.com/OpenBMB/IoA)                                                          | An open-source framework for collaborative AI agents, enabling diverse, distributed agents to team up and tackle complex tasks through internet-like connectivity.                      |  |              |   Agent   |\n| [Jaaz](https://github.com/11cafe/jaaz)                                                         | Jaaz - The world's first open-source multimodal creative assistant. AI design agent, local alternative for Lovart. Canva + Cursor. AI agent with ability to design, edit and generate images, posters, storyboards, etc.                |  |              |   Agent   |\n| [KwaiAgents](https://github.com/KwaiKEG/KwaiAgents)                                            | A generalized information-seeking agent system with Large Language Models (LLMs).                                                     |[arXiv](https://arxiv.org/abs/2312.04889)  |              |   Agent  |\n| [LangChain](https://github.com/langchain-ai/langchain)                                         | Get your LLM application from prototype to production.                                                                                                                                  |        |              |   Agent  |\n| [Langflow](https://github.com/logspace-ai/langflow)                                            | Langflow is a UI for LangChain, designed with react-flow to provide an effortless way to experiment and prototype flows.                                                               |        |              |   Agent  |\n| [LangGraph Studio](https://github.com/langchain-ai/langgraph-studio)                           | LangGraph Studio offers a new way to develop LLM applications by providing a specialized agent IDE that enables visualization, interaction, and debugging of complex agentic applications.        |        |              |   Agent  |\n| [LARP](https://github.com/MiAO-AI-Lab/LARP)                                                    | Language-Agent Role Play for open-world games.                                                                                  |[arXiv](https://arxiv.org/abs/2312.17653)  |              |   Agent  |\n| [LLama Agentic System](https://github.com/meta-llama/llama-agentic-system)                     | Agentic components of the Llama Stack APIs.                                                                                                                                               |      |              |   Agent  |\n| [LlamaIndex](https://github.com/run-llama/llama_index)                                         | LlamaIndex is a data framework for your LLM application.                                                                                                                                  |      |              |   Agent  |\n| [Matrix-Game](https://github.com/SkyworkAI/Matrix-Game)                                        | Matrix-Game: Interactive World Foundation Model. Matrix-Game is a 17B-parameter interactive world foundation model for controllable game world generation.                      |      |              |   Game  |\n| [Matrix-Game 2.0](https://github.com/SkyworkAI/Matrix-Game)                                    | Matrix-Game 2.0: An Open-Source, Real-Time, and Streaming Interactive World Model.                                                                                                        |      |              |   Game  |\n| [MindSearch](https://github.com/InternLM/MindSearch)                                           | 🔍 An LLM-based Multi-agent Framework of Web Search Engine (like Perplexity.ai Pro and SearchGPT).                                                                                     |      |              |   Agent  |\n| [Mixture of Agents (MoA)](https://github.com/togethercomputer/MoA)                             | Mixture-of-Agents Enhances Large Language Model Capabilities.                                                                   |[arXiv](https://arxiv.org/abs/2406.04692)  |              |   Agent  |\n| [MMRole](https://github.com/YanqiDai/MMRole)                                                   | MMRole: A Comprehensive Framework for Developing and Evaluating Multimodal Role-Playing Agents.                               |[arXiv](https://arxiv.org/abs/2408.04203v1)  |              |   Agent  |\n| [Moonlander.ai](https://www.moonlander.ai/)                                                    | Start building 3D games without any coding using generative AI.                                                                                                                          |       |              | Framework |\n| [MuG Diffusion](https://github.com/Keytoyze/Mug-Diffusion)                                     | MuG Diffusion is a charting AI for rhythm games based on Stable Diffusion (one of the most powerful AIGC models) with a large modification to incorporate audio waves.               |       |              |   Game   |\n| [NVIDIA NeMo Agent Toolkit](https://github.com/NVIDIA/NeMo-Agent-Toolkit)                      | NVIDIA NeMo Agent toolkit is a flexible, lightweight, and unifying library that allows you to easily connect existing enterprise agents to data sources and tools across any framework.               |       |              |   Agent   |\n| [Oasis](https://github.com/etched-ai/open-oasis)                                               | Oasis is an interactive world model developed by Decart and Etched. Based on diffusion transformers, Oasis takes in user keyboard input and generates gameplay in an autoregressive manner.                |       |              |   Game   |\n| [OmAgent](https://github.com/om-ai-lab/OmAgent)                                                | A multimodal agent framework for solving complex tasks.                                                                                                                                 |        |              |   Agent  |\n| [OpenAgents](https://github.com/xlang-ai/OpenAgents)                                           | An Open Platform for Language Agents in the Wild.                                                                                                                                       |        |              |   Agent  |\n| [OpenGame](https://github.com/leigest519/OpenGame)                                             | OpenGame: Open Agentic Coding for Games.                               |[arXiv](https://arxiv.org/abs/2604.18394)  |              |   Game  |\n| [Opus](https://opus.ai/)                                                                       | An AI app that turns text into a video game.                                                                                                                                             |       |              |   Game   |\n| [Pipecat](https://github.com/pipecat-ai/pipecat)                                            | Open Source framework for voice and multimodal conversational AI.                                                                                                                           |       |              |   Agent   |\n| [Qwen-Agent](https://github.com/QwenLM/Qwen-Agent)                                             | Qwen-Agent is a framework for developing LLM applications based on the instruction following, tool usage, planning, and memory capabilities of Qwen.                             |        |              |   Agent  |\n| [Ragas](https://github.com/explodinggradients/ragas)                                           | Ragas is a framework that helps you evaluate your Retrieval Augmented Generation (RAG) pipelines.                                                                                     |       |              |   Agent  |\n| [RPBench-Auto](https://github.com/boson-ai/RPBench-Auto)                                       | An automated pipeline for evaluating LLMs for role-playing.                                                                                                                              |       |              |   Game   |\n| [Rosebud AI](https://rosebud.ai)                                                               | Vibe coding platform for creating 3D games and interactive web apps with AI.                                                                                                             |       |              |   Game   |\n| [SIMA](https://deepmind.google/discover/blog/sima-generalist-ai-agent-for-3d-virtual-environments/)          | A generalist AI agent for 3D virtual environments.                                                                                                                         |       |              |   Agent  |\n| [StoryGames.ai](https://storygames.buildbox.com/)                                              | AI for Dreamers Make Games.                                                                                                                                                              |       |              |   Game   |\n| [SWE-agent](https://github.com/princeton-nlp/SWE-agent)                                        | Agent Computer Interfaces Enable Software Engineering Language Models.                                                                |[arXiv](https://arxiv.org/abs/2405.15793)  |              |   Agent  |\n| [TaskGen](https://github.com/simbianai/taskgen)                                                | A Task-based agentic framework building on StrictJSON outputs by LLM agents.                                                                                              |       |              |   Agent  |\n| [TEN Agent](https://github.com/TEN-framework/TEN-Agent)                                        | TEN Agent is the world’s first real-time multimodal agent integrated with the OpenAI Realtime API, RTC, and features weather checks, web search, vision, and RAG capabilities.              |       |              |   Agent  |\n| [Translation Agent](https://github.com/andrewyng/translation-agent)                            | Agentic translation using reflection workflow.                                                                                                                            |       |              |   Agent  |\n| [Twitter](https://github.com/wordware-ai/twitter)                                              | Twitter Personality is a web application that analyzes your Twitter handle to create a personalized personality profile using Wordware AI Agent.                                      |       |              |   Agent  |\n| [Unbounded](https://generative-infinite-game.github.io/)                                         | Unbounded: A Generative Infinite Game of Character Life Simulation.                                                                 |[arXiv](https://arxiv.org/abs/2410.18975)  |              |   Game   |\n| [Video2Game](https://github.com/video2game/video2game)                                         | Real-time, Interactive, Realistic and Browser-Compatible Environment from a Single Video.                                             |[arXiv](https://arxiv.org/abs/2404.09833)  |              |   Game   |\n| [V-IRL](https://virl-platform.github.io/)                                                      | Grounding Virtual Intelligence in Real Life.                                                                                          |[arXiv](https://arxiv.org/abs/2402.03310)  |              |   Agent  |\n| [WebDesignAgent](https://github.com/DAMO-NLP-SG/WebDesignAgent)                                | An agent used for webdesign.                                                                                                                                             |        |              |   Agent  |\n| [XAgent](https://github.com/OpenBMB/XAgent)                                                    | An Autonomous LLM Agent for Complex Task Solving.                                                                                                                                        |       |              |   Agent  |\n\n\u003cp style=\"text-align: right;\"\u003e\u003ca href=\"#table-of-contents\"\u003e^ Back to Contents ^\u003c/a\u003e\u003c/p\u003e\n\n\n## \u003cspan id=\"code\"\u003eCode\u003c/span\u003e\n\n| Source                                                                                      | Description                                                                                                                                                                                    |   Paper   |  Game Engine  |   Type   |\n| :------------------------------------------------------------------------------------------ | :--------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | :-----------: | :-----------: | :-------: |\n| [AI Code Translator](https://github.com/mckaywrigley/ai-code-translator)                       | Use AI to translate code from one language to another.                                                                                                                         |  |        |   Code   |\n| [aiXcoder-7B](https://github.com/aixcoder-plugin/aiXcoder-7B)                                  | aiXcoder-7B Code Large Language Model.                                                                                                           |                                                 |              |   Code   |\n| [bloop](https://github.com/BloopAI/bloop)                                                      | bloop is a fast code search engine written in Rust.                                                                                              |                                                 |              |   Code   |\n| [Chapyter](https://github.com/chapyter/chapyter)                                               | ChatGPT Code Interpreter in Jupyter Notebooks.                                                                                                     |                                               |              |   Code   |\n| [CodeGeeX](https://github.com/THUDM/CodeGeeX)                                                  | An Open Multilingual Code Generation Model.                                                                                   |[arXiv](https://arxiv.org/abs/2303.17568)    |              |   Code   |\n| [CodeGeeX2](https://github.com/THUDM/CodeGeeX2)                                                | A More Powerful Multilingual Code Generation Model.                                                                                               |                                                |              |   Code   |\n| [CodeGeeX4](https://github.com/THUDM/CodeGeeX4)                                                | CodeGeeX4: Open Multilingual Code Generation Model.                                                                                               |                                                |              |   Code   |\n| [CodeGen](https://github.com/salesforce/CodeGen)                                               | CodeGen is an open-source model for program synthesis. Trained on TPU-v4. Competitive with OpenAI Codex.                  |[arXiv](https://arxiv.org/abs/2203.13474)    |              |   Code   |\n| [CodeGen2](https://github.com/salesforce/CodeGen2)                                             | CodeGen2 models for program synthesis.                                                                                        |[arXiv](https://arxiv.org/abs/2305.02309)    |              |   Code   |\n| [Code Llama](https://github.com/facebookresearch/codellama)                                    | Code Llama is a large language models for code based on Llama 2.                                                                                    |                                              |              |   Code   |\n| [CodeTF](https://github.com/salesforce/codetf)                                                 | One-stop Transformer Library for State-of-the-art Code LLM.                                                                                        |                                               |              |   Code   |\n| [CodeT5](https://github.com/salesforce/codet5)                                                 | Open Code LLMs for Code Understanding and Generation.                                                                                              |                                               |              |   Code   |\n| [Code World Model (CWM)](https://github.com/facebookresearch/cwm)                              | Code World Model (CWM) is a 32-billion-parameter open-weights LLM, to advance research on code generation with world models.                       |                                               |              |   Code   |\n| [Cursor](https://www.cursor.so/)                                                               | Write, edit, and chat about your code with GPT-4 in a new type of editor.                                                                          |                                               |              |   Code   |\n| [DeepSeek Coder](https://github.com/deepseek-ai/DeepSeek-Coder)                                | DeepSeek Coder: Let the Code Write Itself.                                                                                    |[arXiv](https://arxiv.org/abs/2401.14196)    |              |   Code   |\n| [OpenAI Codex](https://openai.com/blog/openai-codex)                                           | OpenAI Codex is a descendant of GPT-3.                                                                                                            |                                                |              |   Code   |\n| [PandasAI](https://github.com/gventuri/pandas-ai)                                              | Pandas AI is a Python library that integrates generative artificial intelligence capabilities into Pandas, making dataframes conversational.        |                                     |              |   Code   |\n| [RobloxScripterAI](https://www.haddock.ai/search?platform=Roblox)                              | RobloxScripterAI is an AI-powered code generation tool for Roblox.                                                                                        |                                        |     Roblox    |   Code   |\n| [Roblox GUI Maker](https://robloxguimaker.dev/)                                                | Roblox GUI Maker generates Roblox Studio GUI layouts and Lua starter code from prompts for faster game UI prototyping.                                |                                        |     Roblox    |   Code   |\n| [Scikit-LLM](https://github.com/iryna-kondr/scikit-llm)                                        | Seamlessly integrate powerful language models like ChatGPT into scikit-learn for enhanced text analysis tasks.                                         |                                           |              |   Code   |\n| [SoTaNa](https://github.com/DeepSoftwareAnalytics/SoTaNa)                                      | The Open-Source Software Development Assistant.                                                                               |[arXiv](https://arxiv.org/abs/2308.13416)    |              |   Code   |\n| [Stable Code 3B](https://bit.ly/3O4oGWW)                                                       | Coding on the Edge.                                                                                                                                |                                               |              |   Code   |\n| [StarCoder](https://github.com/bigcode-project/starcoder)                                      | 💫 StarCoder is a language model (LM) trained on source code and natural language text.                                      |[arXiv](https://arxiv.org/abs/2305.06161)    |              |   Code   |\n| [StarCoder 2](https://github.com/bigcode-project/starcoder2)                                   | StarCoder2 is a family of code generation models (3B, 7B, and 15B), trained on 600+ programming languages from The Stack v2 and some natural language text such as Wikipedia, Arxiv, and GitHub issues.   |[arXiv](https://arxiv.org/abs/2402.19173)    |              |   Code   |\n| [Tura](https://github.com/Tura-AI/tura)                                                     | A terminal-native coding agent that turns intent into verified code changes with repo-aware controls and auditable execution.                                             |                                                |              |   Code   |\n| [UnityGen AI](https://github.com/himanshuskyrockets/UnityGen-AI)                               | UnityGen AI is an AI-powered code generation plugin for Unity.                                                                                                 |                                   |     Unity     |   Code   |\n| [Void](https://github.com/voideditor/void)                                                     | Void is an open source Cursor alternative. Write code with the best AI tools, retain full control over your data, and access powerful AI features.             |                                               |              |   Code   |\n\n\u003cp style=\"text-align: right;\"\u003e\u003ca href=\"#table-of-contents\"\u003e^ Back to Contents ^\u003c/a\u003e\u003c/p\u003e\n\n\n## \u003cspan id=\"image\"\u003eImage\u003c/span\u003e\n\n| Source                                                                                      | Description                                                                                                                                                                                    |   Paper   |  Game Engine  |   Type   |\n| :------------------------------------------------------------------------------------------ | :--------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- | :-----------: | :-----------: | :-------: |\n| [AnyDoor](https://ali-vilab.github.io/AnyDoor-Page/)                                           | Zero-shot Object-level Image Customization.                                                                                     |[arXiv](https://arxiv.org/abs/2307.09481)  |              |   Image   |\n| [AnyText](https://github.com/tyxsspa/AnyText)                                                  | Multilingual Visual Text Generation And Editing.                                                                                |[arXiv](https://arxiv.org/abs/2311.03054)  |              |   Image   |\n| [AutoStudio](https://github.com/donahowe/AutoStudio)                                           | Crafting Consistent Subjects in Multi-turn Interactive Image Generation.                                                        |[arXiv](https://arxiv.org/abs/2406.01388)  |              |   Image   |\n| [BAGEL](https://github.com/ByteDance-Seed/Bagel)                                               | BAGEL - Unified Model for Multimodal Understanding and Generation. BAGEL is an open‑source multimodal foundation model with 7B active parameters (14B total) trained on large‑scale interleaved multimodal data.                                                                                                  |[arXiv](https://arxiv.org/abs/2505.14683)  |              |   Image   |\n| [Blender-ControlNet](https://github.com/coolzilj/Blender-ControlNet)                           | Using ControlNet right in Blender.                                                                                              |                                          |    Blender    |   Image   |\n| [BriVL](https://github.com/BAAI-WuDao/BriVL)                                                   | Bridging Vision and Language Model.                                                                                             |[arXiv](https://arxiv.org/abs/2103.06561)  |              |   Image   |\n| [CatVTON](https://github.com/Zheng-Chong/CatVTON)                        ","projects_url":"https://awesome.ecosyste.ms/api/v1/lists/yuan-manx%2Fai-game-devtools/projects"}