Projects in Awesome Lists tagged with mistral
A curated list of projects in awesome lists tagged with mistral .
https://github.com/hiyouga/llama-factory
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
agent ai chatglm fine-tuning gpt instruction-tuning language-model large-language-models llama llama3 llm lora mistral moe peft qlora quantization qwen rlhf transformers
Last synced: 02 Jan 2026
https://github.com/hiyouga/LLaMA-Factory
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
agent ai chatglm fine-tuning gpt instruction-tuning language-model large-language-models llama llama3 llm lora mistral moe peft qlora quantization qwen rlhf transformers
Last synced: 14 Mar 2025
https://github.com/mudler/localai
:robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video, Images, Voice Cloning, Distributed, P2P inference
ai api audio-generation distributed gemma gpt4all image-generation kubernetes libp2p llama llama3 llm mamba mistral musicgen rerank rwkv stable-diffusion text-generation tts
Last synced: 14 May 2026
https://github.com/go-skynet/LocalAI
:robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video, Images, Voice Cloning, Distributed, P2P inference
ai api audio-generation distributed gemma gpt4all image-generation kubernetes libp2p llama llama3 llm mamba mistral musicgen rerank rwkv stable-diffusion text-generation tts
Last synced: 03 May 2025
https://github.com/mudler/LocalAI
:robot: The free, Open Source alternative to OpenAI, Claude and others. Self-hosted and local-first. Drop-in replacement for OpenAI, running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. Features: Generate Text, Audio, Video, Images, Voice Cloning, Distributed inference
ai api audio-generation distributed gemma gpt4all image-generation kubernetes llama llama3 llm mamba mistral musicgen p2p rerank rwkv stable-diffusion text-generation tts
Last synced: 14 Mar 2025
https://github.com/ludwig-ai/ludwig
Low-code framework for building custom LLMs, neural networks, and other AI models
computer-vision data-centric data-science deep deep-learning deeplearning fine-tuning learning llama llama2 llm llm-training machine-learning machinelearning mistral ml natural-language natural-language-processing neural-network pytorch
Last synced: 04 Apr 2026
https://github.com/bentoml/openllm
Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.
bentoml fine-tuning llama llama2 llama3-1 llama3-2 llama3-2-vision llm llm-inference llm-ops llm-serving llmops mistral mlops model-inference open-source-llm openllm vicuna
Last synced: 23 Oct 2025
https://github.com/bentoml/OpenLLM
Run any open-source LLMs, such as Llama 3.1, Gemma, as OpenAI compatible API endpoint in the cloud.
bentoml fine-tuning llama llama2 llama3-1 llama3-2 llama3-2-vision llm llm-inference llm-ops llm-serving llmops mistral mlops model-inference open-source-llm openllm vicuna
Last synced: 14 Mar 2025
https://github.com/xorbitsai/inference
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.
artificial-intelligence chatglm deployment flan-t5 gemma ggml glm4 inference llama llama3 llamacpp llm machine-learning mistral openai-api pytorch qwen vllm whisper wizardlm
Last synced: 25 Apr 2026
https://github.com/yangjianxin1/firefly
Firefly: 大模型训练工具,支持训练Qwen2.5、Qwen2、Yi1.5、Phi-3、Llama3、Gemma、MiniCPM、Yi、Deepseek、Orion、Xverse、Mixtral-8x7B、Zephyr、Mistral、Baichuan2、Llma2、Llama、Qwen、Baichuan、ChatGLM2、InternLM、Ziya2、Vicuna、Bloom等大模型
alpaca aquila baichuan chatglm gemma gpt internlm llama llama2 llama3 llm lora minicpm mistral mixtral peft qlora qwen qwen2 zephyr
Last synced: 14 May 2025
https://github.com/enricoros/big-agi
AI suite powered by state-of-the-art models and providing advanced AI/AGI functions. It features AI personas, AGI functions, multi-model chats, text-to-image, voice, response streaming, code highlighting and execution, PDF import, presets for developers, much more. Deploy on-prem or in the cloud.
agi anthropic beam chatgpt chatgpt-ui generative-ai gpt gpt-4 gpt-5 groq large-language-models mistral multimodal openai openai-api stable-diffusion ui
Last synced: 12 May 2025
https://github.com/yangjianxin1/Firefly
Firefly: 大模型训练工具,支持训练Qwen2.5、Qwen2、Yi1.5、Phi-3、Llama3、Gemma、MiniCPM、Yi、Deepseek、Orion、Xverse、Mixtral-8x7B、Zephyr、Mistral、Baichuan2、Llma2、Llama、Qwen、Baichuan、ChatGLM2、InternLM、Ziya2、Vicuna、Bloom等大模型
alpaca aquila baichuan chatglm gemma gpt internlm llama llama2 llama3 llm lora minicpm mistral mixtral peft qlora qwen qwen2 zephyr
Last synced: 19 Mar 2025
https://linkedin.github.io/Liger-Kernel/
Efficient Triton Kernels for LLM Training
finetuning gemma2 hacktoberfest llama llama3 llm-training llms mistral phi3 triton triton-kernels
Last synced: 12 Feb 2026
https://github.com/gluonfield/enchanted
Enchanted is iOS and macOS app for chatting with private self hosted language models such as Llama2, Mistral or Vicuna using Ollama.
ios large-language-model llama llama2 llm mistral ollama ollama-app swift
Last synced: 13 May 2025
https://github.com/linkedin/liger-kernel
Efficient Triton Kernels for LLM Training
finetuning gemma2 llama llama3 llm-training llms mistral phi3 triton triton-kernels
Last synced: 13 May 2025
https://github.com/mangiucugna/json_repair
A python module to repair invalid JSON from LLMs
deep-learning gpt-4 json llama3 llm machine-learning mistral openai-api parser repair
Last synced: 28 Feb 2026
https://github.com/learningcircuit/local-deep-research
Local Deep Research achieves ~95% on SimpleQA benchmark (tested with GPT-4.1-mini). Supports local and cloud LLMs (Ollama, Google, Anthropic, ...). Searches 10+ sources - arXiv, PubMed, web, and your private documents. Everything Local & Encrypted.
academia anthropic arxiv brave deep-research encryption home-automation homeserver local local-deep-research local-llm mistral ollama openai pubmed research research-tool retrieval-augmented-generation searxng self-hosted
Last synced: 01 May 2026
https://github.com/agentops-ai/agentops
Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including OpenAI Agents SDK, CrewAI, Langchain, Autogen, AG2, and CamelAI
agent agentops agents-sdk ai anthropic autogen cost-estimation crewai evals evaluation-metrics groq langchain llm mistral ollama openai openai-agents
Last synced: 17 Nov 2025
https://github.com/AgentOps-AI/agentops
Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including OpenAI Agents SDK, CrewAI, Langchain, Autogen, AG2, and CamelAI
agent agentops agents-sdk ai anthropic autogen cost-estimation crewai evals evaluation-metrics groq langchain llm mistral ollama openai openai-agents
Last synced: 26 Mar 2025
https://github.com/enricoros/big-AGI
Generative AI suite powered by state-of-the-art models and providing advanced AI/AGI functions. It features AI personas, AGI functions, multi-model chats, text-to-image, voice, response streaming, code highlighting and execution, PDF import, presets for developers, much more. Deploy on-prem or in the cloud.
agi anthropic beam chatgpt chatgpt-ui generative-ai gpt gpt-4 gpt-5 groq large-language-models mistral multimodal openai openai-api stable-diffusion ui
Last synced: 14 Mar 2025
https://github.com/linkedin/Liger-Kernel
Efficient Triton Kernels for LLM Training
finetuning gemma2 llama llama3 llm-training llms mistral phi3 triton triton-kernels
Last synced: 21 Aug 2025
https://github.com/clusterzx/paperless-ai
An automated document analyzer for Paperless-ngx using OpenAI API, Ollama, Deepseek-r1, Azure and all OpenAI API compatible Services to automatically analyze and tag your documents.
ai automation gemma gemma2 llama mistral ollama paperless paperless-ng paperless-ngx phi
Last synced: 14 May 2025
https://github.com/silasmarvin/lsp-ai
LSP-AI is an open-source language server that serves as a backend for AI-powered functionality, designed to assist and empower software engineers, not replace them.
ai auto-completion developer-tools ide language-client llama llamacpp llm lsp mistral openai self-hosted
Last synced: 13 May 2025
https://github.com/stochasticai/xturing
Build, customize and control you own LLMs. From data pre-processing to fine-tuning, xTuring provides an easy way to personalize open-source LLMs. Join our discord community: https://discord.gg/TgHXuSJEk6
adapter alpaca deep-learning fine-tuning finetuning gen-ai generative-ai gpt-2 gpt-j language-model llama llm lora mistral mixed-precision peft quantization
Last synced: 15 May 2025
https://github.com/SilasMarvin/lsp-ai
LSP-AI is an open-source language server that serves as a backend for AI-powered functionality, designed to assist and empower software engineers, not replace them.
ai auto-completion developer-tools ide language-client llama llamacpp llm lsp mistral openai self-hosted
Last synced: 26 Mar 2025
https://github.com/stochasticai/xTuring
Build, customize and control you own LLMs. From data pre-processing to fine-tuning, xTuring provides an easy way to personalize open-source LLMs. Join our discord community: https://discord.gg/TgHXuSJEk6
adapter alpaca deep-learning fine-tuning finetuning gen-ai generative-ai gpt-2 gpt-j language-model llama llm lora mistral mixed-precision peft quantization
Last synced: 13 Mar 2025
https://github.com/floneum/kalosm
Instant, controllable, local pre-trained AI models in Rust
ai candle constrained-generation dioxus floneum-v3 kalosm llama llamacpp llm mistral rust transcription whisper
Last synced: 30 May 2026
https://github.com/darrenburns/elia
A snappy, keyboard-centric terminal user interface for interacting with large language models. Chat with ChatGPT, Claude, Llama 3, Phi 3, Mistral, Gemma and more.
ai chatgpt claude gemma gpt large-language-models llama llama3 llm mistral mistral-ai mixtral ollama ollama-client ollama-interface phi-3 python terminal tui
Last synced: 14 May 2025
https://github.com/lemonade-sdk/lemonade
Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk
ai amd genai gpu llama llm llm-inference local-server mcp mcp-server mistral npu onnxruntime openai-api qwen radeon rocm ryzen vulkan
Last synced: 02 Apr 2026
https://github.com/papersgpt/papersgpt-for-zotero
A powerful Zotero AI and MCP plugin with ChatGPT, Gemini 3, Claude, Grok, DeepSeek, OpenRouter, Kimi, GLM, SiliconFlow, GPT-oss, Gemma 3, Qwen 3
ai chat chatgpt claude deepresearch deepseek gemini gemma3 gpt-5 gpt-oss grok4 kimi llama mcp mistral openrouter pdf qwen3 siliconflow zotero
Last synced: 22 Jan 2026
https://github.com/mobile-artificial-intelligence/maid
Maid is a cross-platform Flutter app for interfacing with GGUF / llama.cpp models locally, and with Ollama and OpenAI models remotely.
android android-ai chatbot chatgpt facebook flutter free-chatgpt gguf large-language-models llama llama-cpp llama2 llamacpp local-ai mistral mobile-ai mobile-artificial-intelligence ollama openai openorca
Last synced: 11 Apr 2025
https://github.com/vitoplantamura/OnnxStream
Lightweight inference library for ONNX files, written in C++. It can run Stable Diffusion XL 1.0 on a RPI Zero 2 (or in 298MB of RAM) but also Mistral 7B on desktops and servers. ARM, x86, WASM, RISC-V supported. Accelerated by XNNPACK.
llama machine-learning mistral onnx raspberry-pi stable-diffusion tinyml wasm webassembly yolov8
Last synced: 17 Apr 2025
https://github.com/vitoplantamura/onnxstream
Lightweight inference library for ONNX files, written in C++. It can run Stable Diffusion XL 1.0 on a RPI Zero 2 (or in 298MB of RAM) but also Mistral 7B on desktops and servers. ARM, x86, WASM, RISC-V supported. Accelerated by XNNPACK.
llama machine-learning mistral onnx raspberry-pi stable-diffusion tinyml wasm webassembly yolov8
Last synced: 14 May 2025
https://github.com/floneum/floneum
Instant, controllable, local pre-trained AI models in Rust
ai candle constrained-generation dioxus floneum-v3 kalosm llama llamacpp llm mistral rust transcription whisper
Last synced: 13 May 2025
https://github.com/Mobile-Artificial-Intelligence/maid
Maid is a cross-platform Flutter app for interfacing with GGUF / llama.cpp models locally, and with Ollama and OpenAI models remotely.
android android-ai chatbot chatgpt facebook flutter free-chatgpt gguf large-language-models llama llama-cpp llama2 llamacpp local-ai mistral mobile-ai mobile-artificial-intelligence ollama openai openorca
Last synced: 24 Mar 2025
https://github.com/n4ze3m/dialoqbase
Create chatbots with ease
anthropic chatbot chatgpt claude code-llama codellama cohere docx google-palm gpt-3 gpt-4 huggingface langchain llama localai mistral pdf private-gpt privategpt
Last synced: 14 May 2025
https://github.com/ai-hypercomputer/maxtext
A simple, performant and scalable Jax LLM!
deepseek fine-tuning gemma2 gemma3 gpt jax large-language-models llama2 llama3 llama4 llm mistral mixtral sft
Last synced: 14 May 2025
https://github.com/zjunlp/EasyEdit
[知识编辑] [ACL 2024] An Easy-to-use Knowledge Editing Framework for LLMs.
artificial-intelligence baichuan chatgpt easyedit efficient gpt knowledge-editing knowlm large-language-models llama llama2 mistral mmedit model-editing natural-language-processing open-source-project safeedit tool trustworthy-ai unlearning
Last synced: 29 Mar 2025
https://github.com/kwaroran/Risuai
Make your own story. User-friendly software for LLM roleplaying
ai characters chat chatbot claude gemini gpt llama llm mcp mcp-client mistral roleplay tauri
Last synced: 22 Apr 2026
https://github.com/vercel/modelfusion
The TypeScript library for building AI applications.
ai artificial-intelligence chatbot claude dall-e embedding gpt-3 huggingface javascript js llamacpp llm mistral multi-modal ollama openai stable-diffusion ts typescript whisper
Last synced: 15 May 2025
https://github.com/CommandCodeAI/BaseAI
BaseAI — The Web AI Framework. The easiest way to build serverless autonomous AI agents with memory. Start building local-first, agentic pipes, tools, and memory. Deploy serverless with one command.
ai anthropic artificial-intelligence baseai cohere firewor gemini grok groq langbase mistral openai perplexity togetherai xai
Last synced: 02 Apr 2026
https://github.com/jakobhoeg/nextjs-ollama-llm-ui
Fully-featured web interface for Ollama LLMs
ai chatbot gemma llm local localstorage mistral mistral-7b nextjs nextjs14 offline ollama openai react shadcn tailwindcss typescript
Last synced: 13 Apr 2025
https://github.com/robitx/gp.nvim
Gp.nvim (GPT prompt) Neovim AI plugin: ChatGPT sessions & Instructable text/code operations & Speech to text [OpenAI, Ollama, Anthropic, ..]
claude codeium copilot gemini gpt-4o gpt4o llm lua mistral neovim nvim ollama parrot perplexity sonnet speech-to-text stt vim voice whisper
Last synced: 14 May 2025
https://github.com/lgrammel/ai-utils.js
The TypeScript library for building AI applications.
ai artificial-intelligence chatbot claude dall-e embedding gpt-3 huggingface javascript js llamacpp llm mistral multi-modal ollama openai stable-diffusion ts typescript whisper
Last synced: 29 Dec 2025
https://github.com/amElnagdy/delegate-skills
Delegate a coding task to a separate coding agent CLI, review the diff, land the commit yourself — one per implementer.
agent-skills ai-coding-agent antigravity claude-code claude-code-skills codex coding-agent cursor grok kimi mistral openai-codex opencode qoder skills-sh
Last synced: 24 Aug 2026
https://github.com/snowby666/poe-api-wrapper
👾 A Python API wrapper for Poe.com. With this, you will have free access to GPT-4, Claude, Llama, Gemini, Mistral and more! 🚀
api chatbot chatgpt claude code-llama dall-e gemini gpt-4 groq llama mistral openai palm2 poe poe-api python quora qwen reverse-engineering stable-diffusion
Last synced: 13 Mar 2025
https://github.com/microsoft/ai-dev-gallery
An open-source project for Windows developers to learn how to add AI with local models and APIs to Windows apps.
ai csharp developer-tools directml dotnet genai mistral npu onnx onnxruntime onnxruntime-genai phi3 qnn stable-diffusion visual-studio whisper winappsdk windows winui3 wpf
Last synced: 14 May 2025
https://github.com/brucemacd/chatd
Chat with your documents using local AI
chat desktop electron llama2 llm mistral mistral-7b ollama rag
Last synced: 09 Apr 2025
https://github.com/BruceMacD/chatd
Chat with your documents using local AI
chat desktop electron llama2 llm mistral mistral-7b ollama rag
Last synced: 06 Apr 2025
https://github.com/aws-samples/generative-ai-use-cases
Application implementation with business use cases for safely utilizing generative AI in business operations
aws bedrock chatbot claude claude3 command-r deepseek-r1 generative-ai image-generation lambda llama3 llm mistral nova rag react sagemaker typescript
Last synced: 17 Nov 2025
https://github.com/hexabot-ai/hexabot
Hexabot v3 is an AI automation platform, combining workflows, actions, agents, and conversational channels in one runtime.
agent agentic agents ai ai-automation artificial-intelligence automation bot-framework chatbot chatgpt claude-ai conversational-ai deepseek framework gemini llama llm mistral ollama workflow
Last synced: 27 Jun 2026
https://github.com/icereed/paperless-gpt
Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI
ai chatgpt llm mistral ocr ollama paperless paperless-ngx
Last synced: 15 May 2025
https://github.com/aws-samples/generative-ai-use-cases-jp
すぐに業務活用できるビジネスユースケース集付きの安全な生成AIアプリ実装
aws bedrock chatbot claude claude3 command-r deepseek-r1 generative-ai image-generation lambda llama3 llm mistral nova rag react sagemaker typescript
Last synced: 30 Mar 2025
https://github.com/open-compass/mixtralkit
A toolkit for inference and evaluation of 'mixtral-8x7b-32kseqlen' from Mistral AI
Last synced: 12 Apr 2025
https://github.com/open-compass/MixtralKit
A toolkit for inference and evaluation of 'mixtral-8x7b-32kseqlen' from Mistral AI
Last synced: 12 Apr 2025
https://github.com/hexastack/hexabot
Hexabot is an open-source AI chatbot / agent builder. It allows you to create and manage multi-channel and multilingual chatbots / agents with ease.
agent agents ai artificial-intelligence bot-framework chatbot chatbot-framework chatbots chatgpt claude-ai conversational-ai gemini grok hacktoberfest llama llm mistral nlu ollama openassistant
Last synced: 15 May 2025
https://github.com/if-ai/comfyui-if_ai_tools
ComfyUI-IF_AI_tools is a set of custom nodes for ComfyUI that allows you to generate prompts using a local Large Language Model (LLM) via Ollama. This tool enables you to enhance your image generation workflow by leveraging the power of language models.
anthropic comfyui flux gemini graphrag groq koboldcpp llamacpp lmstudio mistral ocr ollama omost rag stable-diffusion supervision textgeneration transformers xai
Last synced: 15 May 2025
https://github.com/frankroeder/parrot.nvim
parrot.nvim 🦜 - the plugin that brings stochastic parrots to Neovim.
anthropic chatgpt claude-3-5-sonnet gemini gpt gpt-4o groq-api large-language-models llm mistral neovim nvidia-api nvim o1 ollama openai openai-api perplexity plugin prompting
Last synced: 30 Oct 2025
https://github.com/capsize-games/airunner
Privacy focused, local-first, multi-modal inference engine and agent platform for running LLMs, image generation, speech processing, and tool-based automation
ai ai-art art asset-generator chatbot deep-learning desktop-app image-generation mistral multimodal privacy pygame pyside6 python self-hosted speech-to-text stable-diffusion text-to-image text-to-speech text-to-speech-app
Last synced: 12 Dec 2025
https://github.com/pgalko/bambooai
A Python library powered by Language Models (LLMs) for conversational data discovery and analysis.
ai ai-agents anthropic data-analysis data-science docker gemini groq llm mistral ollama openai-api pandas pinecone python vector-database vllm
Last synced: 15 May 2025
https://github.com/if-ai/ComfyUI-IF_AI_tools
ComfyUI-IF_AI_tools is a set of custom nodes for ComfyUI that allows you to generate prompts using a local Large Language Model (LLM) via Ollama. This tool enables you to enhance your image generation workflow by leveraging the power of language models.
anthropic comfyui flux gemini graphrag groq koboldcpp llamacpp lmstudio mistral ocr ollama omost rag stable-diffusion supervision textgeneration transformers xai
Last synced: 19 Aug 2025
https://github.com/jakobdylanc/llmcord
Make Discord your LLM frontend ● Supports any OpenAI compatible API (Ollama, LM Studio, vLLM, OpenRouter, xAI, Mistral, Groq and more)
bot chat chatbot discord frontend gpt gpt-4 grok groq llama llama3 llama4 llm mistral ollama oobabooga openai vllm xai
Last synced: 15 May 2025
https://github.com/blarc/ai-commits-intellij-plugin
AI Commits for IntelliJ based IDEs/Android Studio.
ai anthropic chatgpt claude commit commit-message gemini generation huggingface intellij intellij-plugin jetbrains llm mistral ollama openai qianfan
Last synced: 15 May 2025
https://github.com/datvodinh/rag-chatbot
Chat with multiple PDFs locally
chatbot chatbot-ui chatbots gradio llama-index llama3 llm mistral ollama question-answering rag
Last synced: 09 Jan 2026
https://github.com/timmyy123/LLM-Hub
Local AI Assistant
ai gemma gemma4 gemma4-agent-skills gptoss granite lfm25 llama llm llm-inference mistral phi4 rag stable-diffusion whisper
Last synced: 12 Aug 2026
https://github.com/evilpsycho/play-with-llms
Tutorial on training, evaluating LLM, as well as utilizing RAG, Agent, Chain to build entertaining applications with LLMs.分享如何训练、评估LLMs,如何基于RAG、Agent、Chain构建有趣的LLMs应用。
agent baichuan2 chatgpt gpt large-language-models llama2 llms mistral rag retrieval-augmented-generation
Last synced: 04 Apr 2025
https://github.com/llm-tools/embedjs
A NodeJS RAG framework to easily work with LLMs and embeddings
ai chatgpt claude cohere embedding embeddings gpt gpt-4 gpt-4o huggingface large-language-models llm mistral ollama openai pinecone rag vector-database vertex-ai
Last synced: 08 Oct 2025
https://github.com/tak-bro/aicommit2
A Reactive CLI that generates commit messages for Git and Jujutsu with Ollama, ChatGPT, Gemini, Claude, Mistral and other AI
ai-commits aicommit aicommits anthropic chatgpt claude cli codestral cohere deepseek git-commit groq jj jujutsu llama mistral ollama perplexity pre-commit pre-commit-hook
Last synced: 10 May 2026
https://github.com/vgel/repeng
A library for making RepE control vectors
language-model machine-learning mistral mistral-7b representation-engineering transformers
Last synced: 21 Aug 2025
https://github.com/sozercan/aikit
🏗️ Fine-tune, build, and deploy open-source LLMs easily!
ai buildkit chatgpt docker fine-tuning finetuning gemma gpt inference kubernetes large-language-models llama llm localllama mistral mixtral nvidia open-llm open-source-llm openai
Last synced: 16 May 2025
https://github.com/pgalko/BambooAI
A lightweight library that leverages Language Models (LLMs) to enable natural language interactions, allowing you to source and converse with data.
ai ai-agents data-analysis data-science gemini groq llm mistral ollama openai-api pandas pinecone python vector-database
Last synced: 23 Mar 2025
https://github.com/devoxx/devoxxgenieideaplugin
DevoxxGenie is a plugin for IntelliJ IDEA that uses local LLM's (Ollama, LMStudio, GPT4All, Jan and Llama.cpp) and Cloud based LLMs to help review, test, explain your project code.
anthropic assistant azure-ai chatgpt chatgpt-api claude-3 claude-ai copilot copilot-chat gemini genai gpt4all groq intellij-plugin java llm lmstudio mistral ollama openai
Last synced: 22 Feb 2026
https://github.com/voxos-ai/bolna
End-to-end platform for building voice first multimodal agents
anyscale chatgpt-api claude-3-sonnet deepgram elevenlabs fastapi gpt-4o llama3 llm mistral openai perplexity-api polly telephony twilio voice-assistant websocket-chat websockets whisper xtts
Last synced: 15 May 2025
https://github.com/bobazooba/xllm
🦖 X—LLM: Cutting Edge & Easy LLM Finetuning
alpaca bitsandbytes cerebras chatgpt deep-learning deep-neural-networks gpt gpt-4 gptq large-language-models llama llama2 llm mistral openai pytorch torch vicuna zephyr
Last synced: 04 Apr 2025
https://github.com/Capsize-Games/airunner
A privacy focused, local-first, multi-modal inference engine and agent platform for running LLMs, image generation, speech processing, and tool-based automation
ai ai-art art asset-generator chatbot deep-learning desktop-app image-generation mistral multimodal privacy pygame pyside6 python self-hosted speech-to-text stable-diffusion text-to-image text-to-speech text-to-speech-app
Last synced: 22 Apr 2025
https://github.com/moritztng/fltr
Like grep but for natural language questions. Based on Mistral 7B or Mixtral 8x7B.
cli grep grep-like llama llama-2 llm localllama mistral mixtral mixtral-8x7b operating-system rust
Last synced: 17 Jan 2026
https://github.com/princeton-nlp/less
[ICML 2024] LESS: Selecting Influential Data for Targeted Instruction Tuning
data data-selection influence instruction-tuning llama llm mistral
Last synced: 05 Apr 2025
https://github.com/SqueezeAILab/KVQuant
[NeurIPS 2024] KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
compression efficient-inference efficient-model large-language-models llama llm localllama localllm mistral model-compression natural-language-processing quantization small-models text-generation transformer
Last synced: 08 May 2025
https://github.com/riccardomusmeci/mlx-llm
Large Language Models (LLMs) applications and tools running on Apple Silicon in real-time with Apple MLX.
llama llm mistral mlx phi transformers
Last synced: 04 Apr 2025
https://github.com/ai-commandos/llama2lang
Convenience scripts to finetune (chat-)LLaMa3 and other models for any language
ai genai huggingface llama2 llama3 llm mistral
Last synced: 05 Apr 2025
https://github.com/llm-tools/embedJs
A NodeJS RAG framework to easily work with LLMs and embeddings
ai chatgpt claude cohere embedding embeddings gpt gpt-4 gpt-4o huggingface large-language-models llm mistral ollama openai pinecone rag vector-database vertex-ai
Last synced: 11 Apr 2025
https://github.com/kevinhermawan/ollamakit
Ollama client for Swift
llama3 llm mistral ollama ollama-api ollama-client
Last synced: 16 May 2025
https://github.com/squeezeailab/kvquant
[NeurIPS 2024] KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
compression efficient-inference efficient-model large-language-models llama llm localllama localllm mistral model-compression natural-language-processing quantization small-models text-generation transformer
Last synced: 07 Apr 2025
https://github.com/codebam/cf-workers-telegram-bot
Telegram Bot library for CloudFlare Workers
ai cloudflare llama2 llama3 mistral telegram telegram-business webhook worker
Last synced: 30 Sep 2025
https://github.com/timmyy123/llm-hub
Local AI Assistant on your phone
ai gemma3 gemma3n gemma4 gemma4-agent-skills gptoss granite lfm25 llama llm llm-inference mistral phi4 rag stable-diffusion
Last synced: 17 Apr 2026
https://github.com/hmunachi/nanodl
A Jax-based library for designing and training transformer models from scratch.
attention attention-mechanism deep-learning distributed-training flax gpt jax llama machine-learning mistral nlp transformer
Last synced: 05 Apr 2025
https://github.com/andrewkchan/yalm
Yet Another Language Model: LLM inference in C++/CUDA, no libraries except for I/O
cpp cuda inference-engine llama llamacpp llm llm-inference machine-learning mistral
Last synced: 12 Apr 2025
https://github.com/jorge-menjivar/unsaged
Open source chat kit engineered for seamless interaction with AI models.
anthropic anthropic-claude bard chatbot chatgpt claude claude-ai claude2 codellama gpt-3 gpt-3-5-turbo gpt-4 gpt-4-turbo llama2 mistral ollama openai palm palm2
Last synced: 09 Apr 2025
https://github.com/gurpreetkaurjethra/end-to-end-generative-ai-projects
End to End Generative AI Industry Projects on LLM Models with Deployment_Awesome LLM Projects
chainlit finetuning-llms gemini generative-ai gpt4o gradio-python-llm huggingface langchain large-language-models llama llama-index llama3 llama3-meta-ai llm llmops lora mergekit mistral openai-api qlora
Last synced: 13 Apr 2025