An open API service indexing awesome lists of open source software.

Projects in Awesome Lists tagged with mistral

A curated list of projects in awesome lists tagged with mistral .

https://github.com/ollama/ollama

Get up and running with Kimi-K2.5, GLM-4.7, DeepSeek, gpt-oss, Qwen, Gemma and other models.

deepseek gemma gemma3 glm go golang gpt-oss llama llama3 llm llms minimax mistral ollama qwen

Last synced: 24 Apr 2026

https://github.com/mudler/localai

:robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video, Images, Voice Cloning, Distributed, P2P inference

ai api audio-generation distributed gemma gpt4all image-generation kubernetes libp2p llama llama3 llm mamba mistral musicgen rerank rwkv stable-diffusion text-generation tts

Last synced: 14 May 2026

https://github.com/go-skynet/LocalAI

:robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video, Images, Voice Cloning, Distributed, P2P inference

ai api audio-generation distributed gemma gpt4all image-generation kubernetes libp2p llama llama3 llm mamba mistral musicgen rerank rwkv stable-diffusion text-generation tts

Last synced: 03 May 2025

https://github.com/mudler/LocalAI

:robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video, Images, Voice Cloning, Distributed inference

ai api audio-generation distributed gemma gpt4all image-generation kubernetes llama llama3 llm mamba mistral musicgen p2p rerank rwkv stable-diffusion text-generation tts

Last synced: 14 Mar 2025

https://github.com/bentoml/openllm

Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.

bentoml fine-tuning llama llama2 llama3-1 llama3-2 llama3-2-vision llm llm-inference llm-ops llm-serving llmops mistral mlops model-inference open-source-llm openllm vicuna

Last synced: 23 Oct 2025

https://github.com/bentoml/OpenLLM

Run any open-source LLMs, such as Llama 3.1, Gemma, as OpenAI compatible API endpoint in the cloud.

bentoml fine-tuning llama llama2 llama3-1 llama3-2 llama3-2-vision llm llm-inference llm-ops llm-serving llmops mistral mlops model-inference open-source-llm openllm vicuna

Last synced: 14 Mar 2025

https://github.com/xorbitsai/inference

Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.

artificial-intelligence chatglm deployment flan-t5 gemma ggml glm4 inference llama llama3 llamacpp llm machine-learning mistral openai-api pytorch qwen vllm whisper wizardlm

Last synced: 25 Apr 2026

https://github.com/LostRuins/koboldcpp

Run GGUF models easily with a KoboldAI UI. One File. Zero Install.

gemma ggml gguf koboldai koboldcpp language-model llama llamacpp llm mistral

Last synced: 23 Mar 2025

https://github.com/yangjianxin1/firefly

Firefly: 大模型训练工具,支持训练Qwen2.5、Qwen2、Yi1.5、Phi-3、Llama3、Gemma、MiniCPM、Yi、Deepseek、Orion、Xverse、Mixtral-8x7B、Zephyr、Mistral、Baichuan2、Llma2、Llama、Qwen、Baichuan、ChatGLM2、InternLM、Ziya2、Vicuna、Bloom等大模型

alpaca aquila baichuan chatglm gemma gpt internlm llama llama2 llama3 llm lora minicpm mistral mixtral peft qlora qwen qwen2 zephyr

Last synced: 14 May 2025

https://github.com/enricoros/big-agi

AI suite powered by state-of-the-art models and providing advanced AI/AGI functions. It features AI personas, AGI functions, multi-model chats, text-to-image, voice, response streaming, code highlighting and execution, PDF import, presets for developers, much more. Deploy on-prem or in the cloud.

agi anthropic beam chatgpt chatgpt-ui generative-ai gpt gpt-4 gpt-5 groq large-language-models mistral multimodal openai openai-api stable-diffusion ui

Last synced: 12 May 2025

https://github.com/yangjianxin1/Firefly

Firefly: 大模型训练工具,支持训练Qwen2.5、Qwen2、Yi1.5、Phi-3、Llama3、Gemma、MiniCPM、Yi、Deepseek、Orion、Xverse、Mixtral-8x7B、Zephyr、Mistral、Baichuan2、Llma2、Llama、Qwen、Baichuan、ChatGLM2、InternLM、Ziya2、Vicuna、Bloom等大模型

alpaca aquila baichuan chatglm gemma gpt internlm llama llama2 llama3 llm lora minicpm mistral mixtral peft qlora qwen qwen2 zephyr

Last synced: 19 Mar 2025

https://github.com/gluonfield/enchanted

Enchanted is iOS and macOS app for chatting with private self hosted language models such as Llama2, Mistral or Vicuna using Ollama.

ios large-language-model llama llama2 llm mistral ollama ollama-app swift

Last synced: 13 May 2025

https://github.com/linkedin/liger-kernel

Efficient Triton Kernels for LLM Training

finetuning gemma2 llama llama3 llm-training llms mistral phi3 triton triton-kernels

Last synced: 13 May 2025

https://github.com/mangiucugna/json_repair

A python module to repair invalid JSON from LLMs

deep-learning gpt-4 json llama3 llm machine-learning mistral openai-api parser repair

Last synced: 28 Feb 2026

https://github.com/learningcircuit/local-deep-research

Local Deep Research achieves ~95% on SimpleQA benchmark (tested with GPT-4.1-mini). Supports local and cloud LLMs (Ollama, Google, Anthropic, ...). Searches 10+ sources - arXiv, PubMed, web, and your private documents. Everything Local & Encrypted.

academia anthropic arxiv brave deep-research encryption home-automation homeserver local local-deep-research local-llm mistral ollama openai pubmed research research-tool retrieval-augmented-generation searxng self-hosted

Last synced: 01 May 2026

https://github.com/agentops-ai/agentops

Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including OpenAI Agents SDK, CrewAI, Langchain, Autogen, AG2, and CamelAI

agent agentops agents-sdk ai anthropic autogen cost-estimation crewai evals evaluation-metrics groq langchain llm mistral ollama openai openai-agents

Last synced: 17 Nov 2025

https://github.com/AgentOps-AI/agentops

Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including OpenAI Agents SDK, CrewAI, Langchain, Autogen, AG2, and CamelAI

agent agentops agents-sdk ai anthropic autogen cost-estimation crewai evals evaluation-metrics groq langchain llm mistral ollama openai openai-agents

Last synced: 26 Mar 2025

https://github.com/enricoros/big-AGI

Generative AI suite powered by state-of-the-art models and providing advanced AI/AGI functions. It features AI personas, AGI functions, multi-model chats, text-to-image, voice, response streaming, code highlighting and execution, PDF import, presets for developers, much more. Deploy on-prem or in the cloud.

agi anthropic beam chatgpt chatgpt-ui generative-ai gpt gpt-4 gpt-5 groq large-language-models mistral multimodal openai openai-api stable-diffusion ui

Last synced: 14 Mar 2025

https://github.com/linkedin/Liger-Kernel

Efficient Triton Kernels for LLM Training

finetuning gemma2 llama llama3 llm-training llms mistral phi3 triton triton-kernels

Last synced: 21 Aug 2025

https://github.com/clusterzx/paperless-ai

An automated document analyzer for Paperless-ngx using OpenAI API, Ollama, Deepseek-r1, Azure and all OpenAI API compatible Services to automatically analyze and tag your documents.

ai automation gemma gemma2 llama mistral ollama paperless paperless-ng paperless-ngx phi

Last synced: 14 May 2025

https://github.com/silasmarvin/lsp-ai

LSP-AI is an open-source language server that serves as a backend for AI-powered functionality, designed to assist and empower software engineers, not replace them.

ai auto-completion developer-tools ide language-client llama llamacpp llm lsp mistral openai self-hosted

Last synced: 13 May 2025

https://github.com/stochasticai/xturing

Build, customize and control you own LLMs. From data pre-processing to fine-tuning, xTuring provides an easy way to personalize open-source LLMs. Join our discord community: https://discord.gg/TgHXuSJEk6

adapter alpaca deep-learning fine-tuning finetuning gen-ai generative-ai gpt-2 gpt-j language-model llama llm lora mistral mixed-precision peft quantization

Last synced: 15 May 2025

https://github.com/SilasMarvin/lsp-ai

LSP-AI is an open-source language server that serves as a backend for AI-powered functionality, designed to assist and empower software engineers, not replace them.

ai auto-completion developer-tools ide language-client llama llamacpp llm lsp mistral openai self-hosted

Last synced: 26 Mar 2025

https://github.com/stochasticai/xTuring

Build, customize and control you own LLMs. From data pre-processing to fine-tuning, xTuring provides an easy way to personalize open-source LLMs. Join our discord community: https://discord.gg/TgHXuSJEk6

adapter alpaca deep-learning fine-tuning finetuning gen-ai generative-ai gpt-2 gpt-j language-model llama llm lora mistral mixed-precision peft quantization

Last synced: 13 Mar 2025

https://github.com/floneum/kalosm

Instant, controllable, local pre-trained AI models in Rust

ai candle constrained-generation dioxus floneum-v3 kalosm llama llamacpp llm mistral rust transcription whisper

Last synced: 30 May 2026

https://github.com/darrenburns/elia

A snappy, keyboard-centric terminal user interface for interacting with large language models. Chat with ChatGPT, Claude, Llama 3, Phi 3, Mistral, Gemma and more.

ai chatgpt claude gemma gpt large-language-models llama llama3 llm mistral mistral-ai mixtral ollama ollama-client ollama-interface phi-3 python terminal tui

Last synced: 14 May 2025

https://github.com/lemonade-sdk/lemonade

Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk

ai amd genai gpu llama llm llm-inference local-server mcp mcp-server mistral npu onnxruntime openai-api qwen radeon rocm ryzen vulkan

Last synced: 02 Apr 2026

https://github.com/papersgpt/papersgpt-for-zotero

A powerful Zotero AI and MCP plugin with ChatGPT, Gemini 3, Claude, Grok, DeepSeek, OpenRouter, Kimi, GLM, SiliconFlow, GPT-oss, Gemma 3, Qwen 3

ai chat chatgpt claude deepresearch deepseek gemini gemma3 gpt-5 gpt-oss grok4 kimi llama mcp mistral openrouter pdf qwen3 siliconflow zotero

Last synced: 22 Jan 2026

https://github.com/mobile-artificial-intelligence/maid

Maid is a cross-platform Flutter app for interfacing with GGUF / llama.cpp models locally, and with Ollama and OpenAI models remotely.

android android-ai chatbot chatgpt facebook flutter free-chatgpt gguf large-language-models llama llama-cpp llama2 llamacpp local-ai mistral mobile-ai mobile-artificial-intelligence ollama openai openorca

Last synced: 11 Apr 2025

https://github.com/vitoplantamura/OnnxStream

Lightweight inference library for ONNX files, written in C++. It can run Stable Diffusion XL 1.0 on a RPI Zero 2 (or in 298MB of RAM) but also Mistral 7B on desktops and servers. ARM, x86, WASM, RISC-V supported. Accelerated by XNNPACK.

llama machine-learning mistral onnx raspberry-pi stable-diffusion tinyml wasm webassembly yolov8

Last synced: 17 Apr 2025

https://github.com/vitoplantamura/onnxstream

Lightweight inference library for ONNX files, written in C++. It can run Stable Diffusion XL 1.0 on a RPI Zero 2 (or in 298MB of RAM) but also Mistral 7B on desktops and servers. ARM, x86, WASM, RISC-V supported. Accelerated by XNNPACK.

llama machine-learning mistral onnx raspberry-pi stable-diffusion tinyml wasm webassembly yolov8

Last synced: 14 May 2025

https://github.com/floneum/floneum

Instant, controllable, local pre-trained AI models in Rust

ai candle constrained-generation dioxus floneum-v3 kalosm llama llamacpp llm mistral rust transcription whisper

Last synced: 13 May 2025

https://github.com/Mobile-Artificial-Intelligence/maid

Maid is a cross-platform Flutter app for interfacing with GGUF / llama.cpp models locally, and with Ollama and OpenAI models remotely.

android android-ai chatbot chatgpt facebook flutter free-chatgpt gguf large-language-models llama llama-cpp llama2 llamacpp local-ai mistral mobile-ai mobile-artificial-intelligence ollama openai openorca

Last synced: 24 Mar 2025

https://github.com/kwaroran/Risuai

Make your own story. User-friendly software for LLM roleplaying

ai characters chat chatbot claude gemini gpt llama llm mcp mcp-client mistral roleplay tauri

Last synced: 22 Apr 2026

https://github.com/CommandCodeAI/BaseAI

BaseAI — The Web AI Framework. The easiest way to build serverless autonomous AI agents with memory. Start building local-first, agentic pipes, tools, and memory. Deploy serverless with one command.

ai anthropic artificial-intelligence baseai cohere firewor gemini grok groq langbase mistral openai perplexity togetherai xai

Last synced: 02 Apr 2026

https://github.com/robitx/gp.nvim

Gp.nvim (GPT prompt) Neovim AI plugin: ChatGPT sessions & Instructable text/code operations & Speech to text [OpenAI, Ollama, Anthropic, ..]

claude codeium copilot gemini gpt-4o gpt4o llm lua mistral neovim nvim ollama parrot perplexity sonnet speech-to-text stt vim voice whisper

Last synced: 14 May 2025

https://github.com/amElnagdy/delegate-skills

Delegate a coding task to a separate coding agent CLI, review the diff, land the commit yourself — one per implementer.

agent-skills ai-coding-agent antigravity claude-code claude-code-skills codex coding-agent cursor grok kimi mistral openai-codex opencode qoder skills-sh

Last synced: 24 Aug 2026

https://github.com/snowby666/poe-api-wrapper

👾 A Python API wrapper for Poe.com. With this, you will have free access to GPT-4, Claude, Llama, Gemini, Mistral and more! 🚀

api chatbot chatgpt claude code-llama dall-e gemini gpt-4 groq llama mistral openai palm2 poe poe-api python quora qwen reverse-engineering stable-diffusion

Last synced: 13 Mar 2025

https://github.com/microsoft/ai-dev-gallery

An open-source project for Windows developers to learn how to add AI with local models and APIs to Windows apps.

ai csharp developer-tools directml dotnet genai mistral npu onnx onnxruntime onnxruntime-genai phi3 qnn stable-diffusion visual-studio whisper winappsdk windows winui3 wpf

Last synced: 14 May 2025

https://github.com/brucemacd/chatd

Chat with your documents using local AI

chat desktop electron llama2 llm mistral mistral-7b ollama rag

Last synced: 09 Apr 2025

https://github.com/BruceMacD/chatd

Chat with your documents using local AI

chat desktop electron llama2 llm mistral mistral-7b ollama rag

Last synced: 06 Apr 2025

https://github.com/kwaroran/risuai

Make your own story. User-friendly software for LLM roleplaying

ai characters chat chatbot claude gemini gpt llama llm mistral roleplay tauri

Last synced: 02 Apr 2026

https://github.com/aws-samples/generative-ai-use-cases

Application implementation with business use cases for safely utilizing generative AI in business operations

aws bedrock chatbot claude claude3 command-r deepseek-r1 generative-ai image-generation lambda llama3 llm mistral nova rag react sagemaker typescript

Last synced: 17 Nov 2025

https://github.com/hexabot-ai/hexabot

Hexabot v3 is an AI automation platform, combining workflows, actions, agents, and conversational channels in one runtime.

agent agentic agents ai ai-automation artificial-intelligence automation bot-framework chatbot chatgpt claude-ai conversational-ai deepseek framework gemini llama llm mistral ollama workflow

Last synced: 27 Jun 2026

https://github.com/kwaroran/RisuAI

Make your own story. User-friendly software for LLM roleplaying

ai characters chat chatbot claude gemini gpt llama llm mistral roleplay tauri

Last synced: 08 Apr 2025

https://github.com/icereed/paperless-gpt

Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI

ai chatgpt llm mistral ocr ollama paperless paperless-ngx

Last synced: 15 May 2025

https://github.com/aws-samples/generative-ai-use-cases-jp

すぐに業務活用できるビジネスユースケース集付きの安全な生成AIアプリ実装

aws bedrock chatbot claude claude3 command-r deepseek-r1 generative-ai image-generation lambda llama3 llm mistral nova rag react sagemaker typescript

Last synced: 30 Mar 2025

https://github.com/open-compass/mixtralkit

A toolkit for inference and evaluation of 'mixtral-8x7b-32kseqlen' from Mistral AI

llm mistral moe

Last synced: 12 Apr 2025

https://github.com/open-compass/MixtralKit

A toolkit for inference and evaluation of 'mixtral-8x7b-32kseqlen' from Mistral AI

llm mistral moe

Last synced: 12 Apr 2025

https://github.com/hexastack/hexabot

Hexabot is an open-source AI chatbot / agent builder. It allows you to create and manage multi-channel and multilingual chatbots / agents with ease.

agent agents ai artificial-intelligence bot-framework chatbot chatbot-framework chatbots chatgpt claude-ai conversational-ai gemini grok hacktoberfest llama llm mistral nlu ollama openassistant

Last synced: 15 May 2025

https://github.com/if-ai/comfyui-if_ai_tools

ComfyUI-IF_AI_tools is a set of custom nodes for ComfyUI that allows you to generate prompts using a local Large Language Model (LLM) via Ollama. This tool enables you to enhance your image generation workflow by leveraging the power of language models.

anthropic comfyui flux gemini graphrag groq koboldcpp llamacpp lmstudio mistral ocr ollama omost rag stable-diffusion supervision textgeneration transformers xai

Last synced: 15 May 2025

https://github.com/capsize-games/airunner

Privacy focused, local-first, multi-modal inference engine and agent platform for running LLMs, image generation, speech processing, and tool-based automation

ai ai-art art asset-generator chatbot deep-learning desktop-app image-generation mistral multimodal privacy pygame pyside6 python self-hosted speech-to-text stable-diffusion text-to-image text-to-speech text-to-speech-app

Last synced: 12 Dec 2025

https://github.com/pgalko/bambooai

A Python library powered by Language Models (LLMs) for conversational data discovery and analysis.

ai ai-agents anthropic data-analysis data-science docker gemini groq llm mistral ollama openai-api pandas pinecone python vector-database vllm

Last synced: 15 May 2025

https://github.com/if-ai/ComfyUI-IF_AI_tools

ComfyUI-IF_AI_tools is a set of custom nodes for ComfyUI that allows you to generate prompts using a local Large Language Model (LLM) via Ollama. This tool enables you to enhance your image generation workflow by leveraging the power of language models.

anthropic comfyui flux gemini graphrag groq koboldcpp llamacpp lmstudio mistral ocr ollama omost rag stable-diffusion supervision textgeneration transformers xai

Last synced: 19 Aug 2025

https://github.com/jakobdylanc/llmcord

Make Discord your LLM frontend ● Supports any OpenAI compatible API (Ollama, LM Studio, vLLM, OpenRouter, xAI, Mistral, Groq and more)

bot chat chatbot discord frontend gpt gpt-4 grok groq llama llama3 llama4 llm mistral ollama oobabooga openai vllm xai

Last synced: 15 May 2025

https://github.com/owlaiproject/owl

A personal wearable AI that runs locally

ai ble bluetooth esp32 llama2 mistral nrf52840 ollama wearable whisper

Last synced: 04 Apr 2025

https://github.com/OwlAIProject/Owl

A personal wearable AI that runs locally

ai ble bluetooth esp32 llama2 mistral nrf52840 ollama wearable whisper

Last synced: 14 Jul 2025

https://github.com/evilpsycho/play-with-llms

Tutorial on training, evaluating LLM, as well as utilizing RAG, Agent, Chain to build entertaining applications with LLMs.分享如何训练、评估LLMs,如何基于RAG、Agent、Chain构建有趣的LLMs应用。

agent baichuan2 chatgpt gpt large-language-models llama2 llms mistral rag retrieval-augmented-generation

Last synced: 04 Apr 2025

https://github.com/tak-bro/aicommit2

A Reactive CLI that generates commit messages for Git and Jujutsu with Ollama, ChatGPT, Gemini, Claude, Mistral and other AI

ai-commits aicommit aicommits anthropic chatgpt claude cli codestral cohere deepseek git-commit groq jj jujutsu llama mistral ollama perplexity pre-commit pre-commit-hook

Last synced: 10 May 2026

https://github.com/pgalko/BambooAI

A lightweight library that leverages Language Models (LLMs) to enable natural language interactions, allowing you to source and converse with data.

ai ai-agents data-analysis data-science gemini groq llm mistral ollama openai-api pandas pinecone python vector-database

Last synced: 23 Mar 2025

https://github.com/devoxx/devoxxgenieideaplugin

DevoxxGenie is a plugin for IntelliJ IDEA that uses local LLM's (Ollama, LMStudio, GPT4All, Jan and Llama.cpp) and Cloud based LLMs to help review, test, explain your project code.

anthropic assistant azure-ai chatgpt chatgpt-api claude-3 claude-ai copilot copilot-chat gemini genai gpt4all groq intellij-plugin java llm lmstudio mistral ollama openai

Last synced: 22 Feb 2026

https://github.com/Capsize-Games/airunner

A privacy focused, local-first, multi-modal inference engine and agent platform for running LLMs, image generation, speech processing, and tool-based automation

ai ai-art art asset-generator chatbot deep-learning desktop-app image-generation mistral multimodal privacy pygame pyside6 python self-hosted speech-to-text stable-diffusion text-to-image text-to-speech text-to-speech-app

Last synced: 22 Apr 2025

https://github.com/moritztng/fltr

Like grep but for natural language questions. Based on Mistral 7B or Mixtral 8x7B.

cli grep grep-like llama llama-2 llm localllama mistral mixtral mixtral-8x7b operating-system rust

Last synced: 17 Jan 2026

https://github.com/princeton-nlp/less

[ICML 2024] LESS: Selecting Influential Data for Targeted Instruction Tuning

data data-selection influence instruction-tuning llama llm mistral

Last synced: 05 Apr 2025

https://github.com/riccardomusmeci/mlx-llm

Large Language Models (LLMs) applications and tools running on Apple Silicon in real-time with Apple MLX.

llama llm mistral mlx phi transformers

Last synced: 04 Apr 2025

https://github.com/ai-commandos/llama2lang

Convenience scripts to finetune (chat-)LLaMa3 and other models for any language

ai genai huggingface llama2 llama3 llm mistral

Last synced: 05 Apr 2025

https://github.com/hmunachi/nanodl

A Jax-based library for designing and training transformer models from scratch.

attention attention-mechanism deep-learning distributed-training flax gpt jax llama machine-learning mistral nlp transformer

Last synced: 05 Apr 2025

https://github.com/andrewkchan/yalm

Yet Another Language Model: LLM inference in C++/CUDA, no libraries except for I/O

cpp cuda inference-engine llama llamacpp llm llm-inference machine-learning mistral

Last synced: 12 Apr 2025

https://github.com/ksylvest/omniai

OmniAI standardizes the APIs for multiple AI providers like OpenAI's Chat GPT, Mistral's LeChat, Claude's Anthropic, Google's Gemini and DeepSeek's Chat..

anthropic chatgpt claude deepseek gemini google lechat mistral omniai openai ruby

Last synced: 02 Apr 2026