An open API service indexing awesome lists of open source software.

awesome-llmops

An awesome & curated list of best LLMOps tools for developers
https://github.com/tensorchord/awesome-llmops

Last synced: 9 days ago
JSON representation

  • AutoML

    • Profiling

      • TPOT - source software packages. | ![GitHub Badge](https://img.shields.io/github/stars/EpistasisLab/tpot.svg?style=flat-square) |
      • Archai - square) |
      • autoai - square) |
      • AutoGL - square) |
      • automl-gs - gs.svg?style=flat-square) |
      • autokeras - team/autokeras.svg?style=flat-square) |
      • Auto-PyTorch - PyTorch.svg?style=flat-square) |
      • auto-sklearn - in replacement for a scikit-learn estimator. | ![GitHub Badge](https://img.shields.io/github/stars/automl/auto-sklearn.svg?style=flat-square) |
      • Dragonfly - square) |
      • Determined - ai/determined.svg?style=flat-square) |
      • DEvol (DeepEvolution) - square) |
      • EvalML - square) |
      • FEDOT - itmo/FEDOT.svg?style=flat-square) |
      • FLAML - us/research/publication/flaml-a-fast-and-lightweight-automl-library/)). | ![GitHub Badge](https://img.shields.io/github/stars/microsoft/FLAML.svg?style=flat-square) |
      • Goptuna - bata/goptuna.svg?style=flat-square) |
      • HpBandSter - square) |
      • Hyperband - square) |
      • Hypernets - square) |
      • Hyperopt - square) |
      • hyperunity - box hyperparameter optimisation. | ![GitHub Badge](https://img.shields.io/github/stars/gdikov/hypertunity.svg?style=flat-square) |
      • Intelli - square) |
      • Katib - native project for automated machine learning (AutoML). | ![GitHub Badge](https://img.shields.io/github/stars/kubeflow/katib.svg?style=flat-square) |
      • Keras Tuner - team/keras-tuner.svg?style=flat-square) |
      • learn2learn - learning Framework for Researchers. | ![GitHub Badge](https://img.shields.io/github/stars/learnables/learn2learn.svg?style=flat-square) |
      • MOE - square) |
      • Model Search - square) |
      • NASGym - of-concept OpenAI Gym environment for Neural Architecture Search (NAS). | ![GitHub Badge](https://img.shields.io/github/stars/gomerudo/nas-env.svg?style=flat-square) |
      • NNI - parameter tuning. | ![GitHub Badge](https://img.shields.io/github/stars/Microsoft/nni.svg?style=flat-square) |
      • Optuna - square) |
      • Pycaret - source, low-code machine learning library in Python that automates machine learning workflows. | ![GitHub Badge](https://img.shields.io/github/stars/pycaret/pycaret.svg?style=flat-square) |
      • REMBO - dimensions via random embedding. | ![GitHub Badge](https://img.shields.io/github/stars/ziyuw/rembo.svg?style=flat-square) |
      • RoBO - square) |
      • scikit-optimize(skopt) - based optimization with a `scipy.optimize` interface. | ![GitHub Badge](https://img.shields.io/github/stars/scikit-optimize/scikit-optimize.svg?style=flat-square) |
      • Spearmint - square) |
      • Torchmeta - Learning library for PyTorch. | ![GitHub Badge](https://img.shields.io/github/stars/tristandeleu/pytorch-meta.svg?style=flat-square) |
      • Vegas - noah/vega.svg?style=flat-square) |
      • AutoRAG - Boost your LLM app performance with your own data | ![GitHub Badge](https://img.shields.io/github/stars/Marker-Inc-Korea/AutoRAG.svg?style=flat-square) |
      • AutoGluon - square) |
      • FEDOT - itmo/FEDOT.svg?style=flat-square) |
      • Ludwig - square) |
  • Awesome Lists

  • Code AI

      • CodeGeeX - square) |
      • CodeGen - source model for program synthesis. Trained on TPU-v4. Competitive with OpenAI Codex. | ![GitHub Badge](https://img.shields.io/github/stars/salesforce/CodeGen.svg?style=flat-square) |
      • CodeT5 - square) |
      • Continue - source autopilot for software development—bring the power of ChatGPT to VS Code | ![GitHub Badge](https://img.shields.io/github/stars/continuedev/continue.svg?style=flat-square) |
      • fauxpilot - source alternative to GitHub Copilot server | ![GitHub Badge](https://img.shields.io/github/stars/fauxpilot/fauxpilot.svg?style=flat-square) |
      • tabby - hosted AI coding assistant. An opensource / on-prem alternative to GitHub Copilot. | ![GitHub Badge](https://img.shields.io/github/stars/TabbyML/tabby.svg?style=flat-square) |
      • promptext - square) |
      • AgentsMesh - hostable AI Agent Workforce Platform. Multi-agent orchestration with remote AI workstations (AgentPods), PTY sandbox + git worktree isolation, built-in Kanban, and per-pod MCP server. Supports Claude Code, Codex CLI, Gemini CLI, Aider, OpenCode. | ![GitHub Badge](https://img.shields.io/github/stars/AgentsMesh/AgentsMesh.svg?style=flat-square) |
      • Bernstein - class MCP server, quality gates, cost tracking with budgets. | ![GitHub Badge](https://img.shields.io/github/stars/sipyourdrink-ltd/bernstein.svg?style=flat-square) |
      • AIDE - source ML engineering agent that uses tree search to explore solution spaces. Automates machine learning experimentation from data analysis to model training. [Paper](https://arxiv.org/abs/2502.13138). | ![GitHub Badge](https://img.shields.io/github/stars/WecoAI/aideml.svg?style=flat-square) |
  • Data

    • Data/Feature enrichment

      • Upgini - to-use features from public and community shared data sources and enriches your training dataset with only the accuracy improving features | ![GitHub Badge](https://img.shields.io/github/stars/upgini/upgini.svg?style=flat-square) |
      • Feast - dev/feast.svg?style=flat-square) |
      • distilabel - quality outputs, full data ownership, and overall efficiency. | ![GitHub Badge](https://img.shields.io/github/stars/argilla-io/distilabel.svg?style=flat-square) |
      • FastDatasets - quality training datasets for Large Language Models. | ![GitHub Badge](https://img.shields.io/github/stars/ZhuLinsen/FastDatasets.svg?style=flat-square) |
    • Data Management

      • ArtiVC - square) |
      • Dolt - square) |
      • Delta-Lake - io/delta.svg?style=flat-square) |
      • Pachyderm - square) |
      • Quilt - organizing data hub for S3. | ![GitHub Badge](https://img.shields.io/github/stars/quiltdata/quilt.svg?style=flat-square) |
      • DVC - Git for Data & Models - ML Experiments Management. | ![GitHub Badge](https://img.shields.io/github/stars/iterative/dvc.svg?style=flat-square) |
    • Data Storage

      • JuiceFS - square) |
      • LakeFS - like capabilities for your object storage. | ![GitHub Badge](https://img.shields.io/github/stars/treeverse/lakeFS.svg?style=flat-square) |
      • Lance - ai/lance.svg?style=flat-square) |
    • Data Tracking

      • Piperider - square) |
      • LUX - org/lux.svg?style=flat-square) |
    • Feature Engineering

  • Federated ML

    • Profiling

      • EasyFL - to-use Federated Learning Platform | ![GitHub Badge](https://img.shields.io/github/stars/EasyFL-AI/EasyFL.svg?style=flat-square) |
      • FATE - square) |
      • FedML - scale cross-silo federated learning, cross-device federated learning on smartphones/IoTs, and research simulation. | ![GitHub Badge](https://img.shields.io/github/stars/FedML-AI/FedML.svg?style=flat-square) |
      • Flower - square) |
      • Harmonia - source project aiming at developing systems/infrastructures and libraries to ease the adoption of federated learning (abbreviated to FL) for researches and production usage. | ![GitHub Badge](https://img.shields.io/github/stars/ailabstw/harmonia.svg?style=flat-square) |
      • TensorFlow Federated - square) |
  • Large Scale Deployment

    • ML Platforms

      • MLRun - square) |
      • Weights & Biases - powered applications, featuring W&B Prompts for LLM execution flow visualization, input and output monitoring, and secure management of prompts and LLM chain configurations. | ![GitHub Badge](https://img.shields.io/github/stars/wandb/wandb.svg?style=flat-square) |
      • Hopsworks - tuning and serving LLMs. Hopsworks includes both a feature store and vector database for RAG. | ![GitHub Badge](https://img.shields.io/github/stars/logicalclocks/hopsworks.svg?style=flat-square) |
      • OpenLLM - tune, serve, deploy, and monitor any LLMs with ease. | ![GitHub Badge](https://img.shields.io/github/stars/bentoml/OpenLLM.svg?style=flat-square) |
      • MLflow - square) |
      • ModelFox - square) |
      • Kserve - square) |
      • Kubeflow - square) |
      • Polyaxon - square) |
      • Primehub - square) |
      • OpenModelZ - click machine learning deployment (LLM, text-to-image and so on) at scale on any cluster (GCP, AWS, Lambda labs, your home lab, or even a single machine). | ![GitHub Badge](https://img.shields.io/github/stars/tensorchord/openmodelz.svg?style=flat-square) |
      • Seldon-core - core.svg?style=flat-square) |
      • Starwhale - tuning. | ![GitHub Badge](https://img.shields.io/github/stars/star-whale/starwhale.svg?style=flat-square) |
      • TrueFoundry - tune and serve LLM Models on a company’s own Infrastructure with Data Security and Optimal GPU and Cost Management. Launch your LLM Application at Production scale with best DevSecOps practices. | |
    • Model Management

      • Comet - ml/comet-examples.svg?style=flat-square) |
      • ModelDB - square) |
      • MLEM - square) |
      • ormb - square) |
    • Scheduling

      • PAI - sourced by Microsoft). | ![GitHub Badge](https://img.shields.io/github/stars/microsoft/pai.svg?style=flat-square) |
      • Kueue - native Job Queueing. | ![GitHub Badge](https://img.shields.io/github/stars/kubernetes-sigs/kueue.svg?style=flat-square) |
      • Slurm - square) |
      • Volcano - sh/volcano.svg?style=flat-square) |
      • Yunikorn - weight, universal resource scheduler for container orchestrator systems. | ![GitHub Badge](https://img.shields.io/github/stars/apache/yunikorn-core.svg?style=flat-square) |
    • Workflow

      • Airflow - square) |
      • ZenML - io/zenml.svg?style=flat-square) |
      • aqueduct - Source Platform for Production Data Science | ![GitHub Badge](https://img.shields.io/github/stars/aqueducthq/aqueduct.svg?style=flat-square) |
      • Argo Workflows - workflows.svg?style=flat-square) |
      • Flyte - native workflow automation platform for complex, mission-critical data and ML processes at scale. | ![GitHub Badge](https://img.shields.io/github/stars/flyteorg/flyte.svg?style=flat-square) |
      • Kubeflow Pipelines - square) |
      • Metaflow - life data science projects with ease! | ![GitHub Badge](https://img.shields.io/github/stars/Netflix/metaflow.svg?style=flat-square) |
      • Ploomber - square) |
      • Prefect - square) |
      • simulate-sdk - grade Voice AI simulation SDK for scenario-driven stress testing of multimodal and agentic systems. | ![GitHub Badge](https://img.shields.io/github/stars/future-agi/simulate-sdk?style=flat-square) |
      • VDP - source unstructured data ETL tool to streamline the end-to-end unstructured data processing pipeline. | ![GitHub Badge](https://img.shields.io/github/stars/instill-ai/vdp.svg?style=flat-square) |
      • Hamilton - inc/hamilton.svg?style=flat-square) |
      • Kitaru - io/kitaru.svg?style=flat-square) |
  • LLMOps

    • Observability

      • Portkey - efficient apps. | |
      • Fiddler AI - production to production. | |
      • TrueFoundry - prem) Infra including deploying, Fine-tuning, tracking Prompts and serving Open Source LLM Models with full Data Security and Optimal GPU Management. Train and Launch your LLM Application at Production scale with best Software Engineering practices. | |
      • Parea AI - controlled enhanced prompt playground. | ![GitHub Badge](https://img.shields.io/github/stars/parea-ai/parea-sdk-py?style=flat-square) |
      • Vellum
      • Izlo
      • Keywords AI
      • Literal AI - modal LLM observability and evaluation platform. Create prompt templates, deploy prompts versions, debug LLM runs, create datasets, run evaluations, monitor LLM metrics and collect human feedback. | |
      • Evidently - source framework to evaluate, test and monitor ML and LLM-powered systems. | ![GitHub Badge](https://img.shields.io/github/stars/evidentlyai/evidently.svg?style=flat-square) |
      • agenta - AI/agenta.svg?style=flat-square) |
      • AI studio - square) |
      • Arize-Phoenix - ai/phoenix.svg?style=flat-square) |
      • BudgetML - square) |
      • deeplake - square) |
      • Dify - source framework aims to enable developers (and even non-developers) to quickly build useful applications based on large language models, ensuring they are visual, operable, and improvable. | ![GitHub Badge](https://img.shields.io/github/stars/langgenius/dify.svg?style=flat-square) |
      • Dstack - effective LLM development in any cloud (AWS, GCP, Azure, Lambda, etc). | ![GitHub Badge](https://img.shields.io/github/stars/dstackai/dstack.svg?style=flat-square) |
      • Glide - Native LLM Routing Engine. Improve LLM app resilience and speed. | ![GitHub Badge](https://img.shields.io/github/stars/einstack/glide.svg?style=flat-square) |
      • GPTCache - square) |
      • Haystack - answering and more. | ![GitHub Badge](https://img.shields.io/github/stars/deepset-ai/haystack.svg?style=flat-square) |
      • Langfuse - square) |
      • LangKit - of-the-box LLM telemetry collection library that extracts features and profiles prompts, responses and metadata about how your LLM is performing over time to find problems at scale. | ![GitHub Badge](https://img.shields.io/github/stars/whylabs/langkit.svg?style=flat-square) |
      • LLMApp - time LLM-enabled data pipelines with few lines of code. | ![GitHub Badge](https://img.shields.io/github/stars/pathwaycom/llm-app.svg?style=flat-square) |
      • LLMFlows - answering systems, and agents. | ![GitHub Badge](https://img.shields.io/github/stars/stoyan-stoyanov/llmflows.svg?style=flat-square) |
      • magentic - powered functionality. | ![GitHub Badge](https://img.shields.io/github/stars/jackmpcollins/magentic.svg?style=flat-square) |
      • Mirascope - fast, efficient development and ensuring quality in LLM-based applications | ![GitHub Badge](https://img.shields.io/github/stars/Mirascope/mirascope.svg?style=flat-square) |
      • OpenLIT - native GenAI and LLM Application Observability tool and provides OpenTelmetry Auto-instrumentation for monitoring LLMs, VectorDBs and Frameworks. It provides valuable insights into token & cost usage, user interaction, and performance related metrics. | ![GitHub Badge](https://img.shields.io/github/stars/dokulabs/doku.svg?style=flat-square) |
      • Pezzo 🕹️ - source LLMOps platform built for developers and teams. In just two lines of code, you can seamlessly troubleshoot your AI operations, collaborate and manage your prompts in one place, and instantly deploy changes to any environment. | ![GitHub Badge](https://img.shields.io/github/stars/pezzolabs/pezzo.svg?style=flat-square) |
      • PromptMage - source tool to simplify the process of creating and managing LLM workflows and prompts as a self-hosted solution. | ![GitHub Badge](https://img.shields.io/github/stars/tsterbak/promptmage.svg?style=flat-square) |
      • prompttools - source tools for testing and experimenting with prompts. The core idea is to enable developers to evaluate prompts using familiar interfaces like code and notebooks. In just a few lines of codes, you can test your prompts and parameters across different models (whether you are using OpenAI, Anthropic, or LLaMA models). You can even evaluate the retrieval accuracy of vector databases. | ![GitHub Badge](https://img.shields.io/github/stars/hegelai/prompttools.svg?style=flat-square) |
      • xTuring - tuning. | ![GitHub Badge](https://img.shields.io/github/stars/stochasticai/xturing.svg?style=flat-square) |
      • Helicone - source LLM observability platform for logging, monitoring, and debugging AI applications. Simple 1-line integration to get started. | ![GitHub Badge](https://img.shields.io/github/stars/helicone/helicone.svg?style=flat-square) |
      • PromptFoundry - foundry/python-sdk.svg?style=flat-square) |
      • gotoHuman - based and agentic workflows. Prompt users to approve actions, select next steps, or review and validate generated results. |
      • GPUStack - source GPU cluster manager for running and managing LLMs | ![GitHub Badge](https://img.shields.io/github/stars/gpustack/gpustack.svg?style=flat-square) |
      • PromptLayer 🍰 - layer-library.svg?style=flat-square) |
      • Opik - ml/opik.svg?style=flat-square) |
      • Laminar - source all-in-one platform for engineering AI products. Traces, Evals, Datasets, Labels. | ![GitHub Badge](https://img.shields.io/github/stars/lmnr-ai/lmnr.svg?style=flat-square) |
      • Dataoorts
      • PromptDX - ai/promptdx.svg?style=flat-square) |
      • systemprompt.io
      • MLflow - source framework for the end-to-end machine learning lifecycle, helping developers track experiments, evaluate models/prompts, deploy models, and add observability with tracing. | ![GitHub Badge](https://img.shields.io/github/stars/mlflow/mlflow.svg?style=flat-square) |
      • Epsilla - in-one platform to create vertical AI agents powered by your private data and knowledge. | |
      • PromptSite - works directly with your local filesystem, ideal for data scientists and engineers to easily integrate into existing LLM workflows | |
      • AgentMark - Safe Markdown-based Agents | ![GitHub Badge](https://img.shields.io/github/stars/Puzzlet-ai/agentmark.svg?style=flat-square) |
      • Cheshire Cat AI - cat-ai/core.svg?style=flat-square) |
      • Lunary - and-play integration into LangChain. | ![GitHub Badge](https://img.shields.io/github/stars/lunary-ai/lunary.svg?style=flat-square) |
      • AI studio - square) |
      • LiteLLM 🚅 - square) |
      • LlamaIndex - square) |
      • langchain - square) |
      • Manag.ai - in-one prompt management and observability platform. Craft, track, and perfect your LLM prompts with ease. | |
      • TensorZero - source framework for building production-grade LLM applications. It unifies an LLM gateway, observability, optimization, evaluations, and experimentation. | ![GitHub Badge](https://img.shields.io/github/stars/tensorzero/tensorzero.svg?style=flat-square) |
      • Manag.ai - in-one prompt management and observability platform. Craft, track, and perfect your LLM prompts with ease. | |
      • Dataoorts
      • Hypersigil - source prompt lifecycle management and gateway with a Web UI. | ![GitHub Badge](https://img.shields.io/github/stars/hypersigilhq/hypersigil.svg?style=flat-square) |
      • Neurolink - provider AI agent framework that unifies 12+ LLM providers (OpenAI, Google, Anthropic, AWS, Azure, Groq, etc.) with workflow orchestration. Production-grade platform for building LLM applications with streaming, tool calling, caching, and enterprise features. Battle-tested at 15M+ requests/month. | ![GitHub Badge](https://img.shields.io/github/stars/juspay/neurolink.svg?style=flat-square) |
      • Weights & Biases (Prompts) - first W&B MLOps platform. Utilize W&B Prompts for visualizing and inspecting LLM execution flow, tracking inputs and outputs, viewing intermediate results, securely managing prompts and LLM chain configurations. | |
      • Embedchain - square) |
      • Roundtable - configuration unified AI assistant management built on the FastMCP framework. Provides seamless integration with Claude, ChatGPT, and other AI assistants through a single MCP interface with session management, logging, and production-ready operations. | ![GitHub Badge](https://img.shields.io/github/stars/askbudi/roundtable.svg?style=flat-square) |
      • Future AGI - agi/ai-evaluation?style=flat-square) |
      • PraisonAI - ready Multi-AI Agents framework with self-reflection. Fastest agent instantiation (3.77μs), 100+ LLM support via LiteLLM, MCP integration, agentic workflows (route/parallel/loop/repeat), built-in memory, Python & JS SDKs. | ![GitHub Badge](https://img.shields.io/github/stars/MervinPraison/PraisonAI.svg?style=flat-square) |
      • Puzzlet AI - Based LLM Engineering Platform. Achieve more from GenAI: Manage, evaluate, and improve your full-stack LLM application - with version control, type-safety, and local development built-in. | |
      • LRM - powered translation via Ollama, validation, and code scanning for unused/missing keys. | ![GitHub Badge](https://img.shields.io/github/stars/nickprotop/LocalizationManager.svg?style=flat-square) |
      • TeamoRouter - 4o, Gemini, DeepSeek, Kimi, MiniMax. Smart routing modes (teamo-best, teamo-balanced, teamo-eco) auto-pick the optimal model. Up to 50% off official prices. 2-second install via skill.md. | |
      • Semantic Cache Router - semantic-cache-and-stateful-routing-system.svg?style=flat-square) |
      • AgentField - source control plane for building and operating AI agents like APIs at scale, with routing, memory, observability, identity, auth, and policy controls. | ![GitHub Badge](https://img.shields.io/github/stars/Agent-Field/agentfield.svg?style=flat-square) |
      • Contexto - hosted context engine for AI agents with persistent conversation memory and recall. Works as a drop-in OpenAI-compatible proxy, OpenClaw plugin, or memory SDK — no code changes required. | ![GitHub Badge](https://img.shields.io/github/stars/ekailabs/contexto.svg?style=flat-square) |
      • Hive - source AI agent framework for building goal-driven, self-improving autonomous agents with auto-generated graphs, evolution loops, and MCP integration. | ![GitHub Badge](https://img.shields.io/github/stars/aden-hive/hive.svg?style=flat-square) |
      • Mengram - source memory infrastructure for AI agents. Provides semantic (entities/facts), episodic (conversations), and procedural (learned behaviors) memory with auto-reflection. Python SDK, JS SDK, MCP server, and REST API. | ![GitHub Badge](https://img.shields.io/github/stars/alibaizhanov/mengram.svg?style=flat-square) |
      • TeamoRouter - 4o, Gemini, DeepSeek, Kimi, MiniMax. Smart routing modes (teamo-best, teamo-balanced, teamo-eco) auto-pick the optimal model. Up to 50% off official prices. 2-second install via skill.md. | |
      • Rhesis - source testing infrastructure for LLM and agentic applications. Collaborative platform enabling teams to define quality metrics, run evaluations, and ship confidently with version control and peer review workflows built for AI engineering. | ![GitHub Badge](https://img.shields.io/github/stars/rhesis-ai/rhesis.svg?style=flat-square) |
      • Statewave - source memory runtime for AI agents. Compiles events into deterministic, provenance-tagged context bundles instead of query-time retrieval. Apache-2.0, self-hostable on Postgres + pgvector. | ![GitHub Badge](https://img.shields.io/github/stars/smaramwbc/statewave.svg?style=flat-square) |
      • SwarmClaw - hosted multi-agent AI runtime with 23+ LLM providers, persistent memory, skills, schedules, sub-agent spawning, and MCP client + server support. Ships as desktop app, CLI, or Docker. | ![GitHub Badge](https://img.shields.io/github/stars/swarmclawai/swarmclaw.svg?style=flat-square) |
      • future-agi - source self-hostable end-to-end agent engineering and optimization platform unifying tracing, evals, simulations, datasets, gateway, and guardrails for LLM and AI agent applications. | ![GitHub Badge](https://img.shields.io/github/stars/future-agi/future-agi?style=flat-square) |
      • PromptDX - ai/promptdx.svg?style=flat-square) |
      • Registry Broker - online/registry-broker.svg?style=flat-square) |
  • ML Platforms

    • TrueFoundry - A PaaS to deploy, Fine-tune and serve LLM Models on a company’s own Infrastructure with Data Security and Optimal GPU and Cost Management. Launch your LLM Application at Production scale with best DevSecOps practices.