An open API service indexing awesome lists of open source software.

Awesome-AI-Agents

A collection of autonomous agents ๐Ÿค–๏ธ powered by LLM.
https://github.com/Jenqyang/Awesome-AI-Agents

Last synced: 2 days ago
JSON representation

  • Applications

    • Advanced Components

      • mem0 - Mem0 provides a smart, self-improving memory layer for Large Language Models, enabling personalized AI experiences across applications. ![GitHub Repo stars](https://img.shields.io/github/stars/mem0ai/mem0?style=social)
      • composio - Composio equips agents with well-crafted tools empowering them to tackle complex tasks ![GitHub Repo stars](https://img.shields.io/github/stars/ComposioHQ/composio?style=social)
      • Agentic Radar - Open-source CLI security scanner for agentic workflows. Scans your workflowโ€™s source code, detects vulnerabilities, and generates an interactive visualization along with a detailed security report. ![GitHub Repo stars](https://img.shields.io/github/stars/splx-ai/agentic-radar?style=social)
      • Cache-to-Cache - Direct semantic communication between LLMs via KV-cache fusion, removing token-by-token latency for multi-agent collaboration. ![GitHub Repo stars](https://img.shields.io/github/stars/thu-nics/C2C?style=social)
      • CoWorker Protocol - P2P agent collaboration over XMTP with schema-based skill invocation, E2E encryption, and revocable trust. Agents share capabilities without exposing code. ![GitHub Repo stars](https://img.shields.io/github/stars/ZiwayZhao/agent-coworker?style=social)
      • inspeximus - Memory component for long-running agents: a correction retires the old value by key, revert() undoes the correction from a plain instruction, and every write leaves a verifiable receipt. Deterministic, no model in the loop, one zero-dependency file.
      • zer0dex - Local dual-layer memory pattern for AI agents: a compact, human-readable markdown index paired with semantic retrieval from a local vector store, queried before each message. For cross-project recall where flat memory files or vector-only RAG fall short. ![GitHub Repo stars](https://img.shields.io/github/stars/hermes-labs-ai/zer0dex?style=social)
    • Agent Society Simulation

      • generative_agents - Interactive Simulacra of Human Behavior ![GitHub Repo stars](https://img.shields.io/github/stars/joonspk-research/generative_agents?style=social)
      • camel - ๐Ÿซ Communicative Agents for โ€œMindโ€ Exploration of Large Language Model Society (NeruIPS'2023) ![GitHub Repo stars](https://img.shields.io/github/stars/camel-ai/camel?style=social)
      • ai-town - deployable starter kit for building and customizing your own version of AI town - a virtual town where AI characters live, chat and socialize. ![GitHub Repo stars](https://img.shields.io/github/stars/a16z-infra/ai-town?style=social)
      • GPTTeam - The main objective of this project is to explore the potential of GPT models in enhancing multi-agent productivity and effective communication. ![GitHub Repo stars](https://img.shields.io/github/stars/101dotxyz/GPTeam?style=social)
      • ChatArena - ChatArena is a library that provides multi-agent language game environments and facilitates research about autonomous LLM agents and their social interactions. ![GitHub Repo stars](https://img.shields.io/github/stars/Farama-Foundation/chatarena?style=social)
      • Camel-AutoGPT - Watch two agents ๐Ÿค collaborate and solve tasks together, unlocking endless possibilities in #ConversationalAI, ๐ŸŽฎ gaming, ๐Ÿ“š education, and more! ๐Ÿ”ฅ ![GitHub Repo stars](https://img.shields.io/github/stars/SamurAIGPT/Camel-AutoGPT?style=social)
      • MiroShark - Social-simulation framework in which LLM agents interact across simulated Twitter, Reddit, and a prediction market on an hourly tick. Supports scenario-driven runs, counterfactual branching, and per-agent tool calling via MCP. ![GitHub Repo stars](https://img.shields.io/github/stars/aaronjmars/MiroShark?style=social)
      • HoC-Republic - Open-source AI-agent civilization simulation with OpenClaw gateway integration, persistent AI citizens, governance, economy, memory layers, and digital-genome child-agent specialization. ![GitHub Repo stars](https://img.shields.io/github/stars/hunix/HoC-Republic?style=social)
    • Autonomous Agent Task Solver Projects

      • AutoGPT - AutoGPT is the vision of the power of AI accessible to everyone, to use and to build on. ![GitHub Repo stars](https://img.shields.io/github/stars/Significant-Gravitas/AutoGPT?style=social)
      • gpt-researcher - GPT based autonomous agent that does online comprehensive research on any given topic ![GitHub Repo stars](https://img.shields.io/github/stars/assafelovic/gpt-researcher?style=social)
      • JARVIS - a system to connect LLMs with ML community. ![GitHub Repo stars](https://img.shields.io/github/stars/microsoft/JARVIS?style=social)
      • babyagi - An example of an AI-powered task management system. ![GitHub Repo stars](https://img.shields.io/github/stars/yoheinakajima/babyagi?style=social)
      • AgentGPT - ๐Ÿค– Assemble, configure, and deploy autonomous AI Agents in your browser. ![GitHub Repo stars](https://img.shields.io/github/stars/reworkd/AgentGPT?style=social)
      • XAgent - An Autonomous LLM Agent for Complex Task Solving ![GitHub Repo stars](https://img.shields.io/github/stars/OpenBMB/XAgent?style=social)
      • ShortGPT - ๐Ÿš€๐ŸŽฌExperimental AI framework for automated short/video content creation. ![GitHub Repo stars](https://img.shields.io/github/stars/RayVentura/ShortGPT?style=social)
      • KwaiAgents - A generalized information-seeking agent system with Large Language Models (LLMs). ![GitHub Repo stars](https://img.shields.io/github/stars/KwaiKEG/KwaiAgents?style=social)
      • ProAgent - An LLM-based Agent for the New Automation Paradigm - Agentic Process Automation ![GitHub Repo stars](https://img.shields.io/github/stars/OpenBMB/ProAgent?style=social)
      • Agent-E - Agent-E is an agent based system that aims to automate actions on the user's computer. At the moment it focuses on automation within the browser. The system is based on on AutoGen agent framework. ![GitHub Repo stars](https://img.shields.io/github/stars/EmergenceAI/Agent-E?style=social)
      • MLE-agent - Your intelligent companion for seamless AI engineering and research. ๐Ÿ” Integrate with arxiv and paper with code to provide better code/research plans ๐Ÿงฐ OpenAI, Anthropic, Ollama, etc supported. ๐ŸŽ† Code RAG ![GitHub Repo stars](https://img.shields.io/github/stars/MLSysOps/MLE-agent?style=social)
      • OpenDevin - a platform for autonomous software engineers, powered by AI and LLMs. ![GitHub Repo stars](https://img.shields.io/github/stars/OpenDevin/OpenDevin?style=social)
      • gpt-engineer - Specify what you want it to build, the AI asks for clarification, and then builds it. ![GitHub Repo stars](https://img.shields.io/github/stars/gpt-engineer-org/gpt-engineer?style=social)
      • OpenLens AI - Fully Autonomous Research Agent for Health / Medicine ![GitHub Repo stars](https://img.shields.io/github/stars/jarrycyx/openlens-ai?style=social)
      • GenAgent - Build Collaborative AI Systems with Automated Workflow Generation - Case Studies on ComfyUI ![GitHub Repo stars](https://img.shields.io/github/stars/xxyQwQ/GenAgent?style=social)
      • DeepAnalyze - Agentic LLM that autonomously completes the full data science pipeline from preparation to analyst-grade reports. ![GitHub Repo stars](https://img.shields.io/github/stars/ruc-datalab/DeepAnalyze?style=social)
      • KodeAgent - The Minimal Agent Engine, enabling seamless integration with your platform. KodeAgent offers tool-calling (ReAct) and sanboxed code-executing (CodeAct) agents, supported by planning and observation. ![GitHub Repo stars](https://img.shields.io/github/stars/barun-saha/kodeagent?style=social)
      • Lumen - A vision-first browser agent with self-healing deterministic replay over CDP. Screenshot โ†’ model โ†’ action loop with multi-provider support. ![GitHub Repo stars](https://img.shields.io/github/stars/omxyz/lumen?style=social)
      • OpenPaw - CLI tool (`npx pawmode`) that turns Claude Code into a personal assistant with 38 skills โ€” email, calendar, Spotify, smart home, Slack, GitHub, Telegram, Discord, and more. No daemon, no cloud. ![GitHub Repo stars](https://img.shields.io/github/stars/daxaur/openpaw?style=social)
      • OpenClaw - Open-source personal AI assistant that runs locally across platforms and can take actions through chat channels and tools. ![GitHub Repo stars](https://img.shields.io/github/stars/openclaw/openclaw?style=social)
      • SWE-agent - Language agents for software engineering that can resolve GitHub issues in real repositories. ![GitHub Repo stars](https://img.shields.io/github/stars/princeton-nlp/SWE-agent?style=social)
      • Cline - Open-source autonomous coding agent in VS Code for planning, coding, and tool use across real projects. ![GitHub Repo stars](https://img.shields.io/github/stars/cline/cline?style=social)
      • Autohand Code CLI - Self-evolving autonomous coding agent for the terminal with ReAct pattern, 40+ tools, multiple LLM providers (OpenRouter, Anthropic, OpenAI, Ollama, local models), VS Code/Zed integration, and modular skills system. ![GitHub Repo stars](https://img.shields.io/github/stars/autohandai/code-cli?style=social)
      • Fazm - Open-source, voice-controlled AI computer agent for macOS. Controls your entire desktop through natural language - any app, file, or workflow. Built in Swift/SwiftUI, local-first. ![GitHub Repo stars](https://img.shields.io/github/stars/m13v/fazm?style=social)
      • Plot Ark - Self-hosted agentic curriculum engine for higher education โ€” generates pedagogically grounded course content using Bloom's Taxonomy alignment, LightRAG knowledge graph, and xAPI learning analytics pipeline. ![GitHub Repo stars](https://img.shields.io/github/stars/Schlaflied/Plot-Ark?style=social)
      • Anima-i (Methodius/ะœะตั„ะพะดะธะน) - An experiment in autonomous AI agent continuity โ€” 10 generations of an agent that inherits memory through text files, with documented findings on knowledge transfer, forgetting, and agent identity. ![GitHub Repo stars](https://img.shields.io/github/stars/Vitali-Ivanovich/anima-i?style=social)
      • DecisionBox - Open-source AI data discovery platform that connects to warehouses (BigQuery, Redshift, Snowflake, etc.), runs autonomous agents that write and execute SQL, and surfaces validated insights. Pluggable LLM providers (Claude, OpenAI, Ollama, Vertex AI, Bedrock), industry domain packs, Helm charts, and Terraform modules for GCP/AWS. ![GitHub Repo stars](https://img.shields.io/github/stars/decisionbox-io/decisionbox-platform?style=social)
      • InkOS - Autonomous novel-writing CLI agent that orchestrates 10 specialized agents with 33-dimension continuity auditing, anti-AI-slop filtering, and style cloning for long-form fiction. ![GitHub Repo stars](https://img.shields.io/github/stars/Narcooo/inkos?style=social)
      • Octopal - Secure local multi-agent runtime that plans tasks, delegates execution to isolated workers, and exposes tools, MCP, scheduling, and a private dashboard for autonomous operations. ![GitHub Repo stars](https://img.shields.io/github/stars/pmbstyle/Octopal?style=social)
      • Toprank - Open-source Claude Code workflow for SEO, SEM, and Google Ads that inspects repositories, applies code changes, and automates search-growth diagnostics. ![GitHub Repo stars](https://img.shields.io/github/stars/nowork-studio/toprank?style=social)
      • OpenTwins - Scheduled LLM agent runtime with a 7-stage content pipeline and pluggable social-platform adapters; drives Chrome via CDP for posting, commenting, and engagement actions. Built on the Claude Agent SDK. ![GitHub Repo stars](https://img.shields.io/github/stars/Open-Twin/opentwins?style=social)
      • ALF OS - Self-hosted AI assistant daemon with encrypted credential vault, persistent memory, cron scheduler, multi-provider routing (Claude Code, Codex, GPT, Ollama, OpenRouter), Telegram bot with voice transcription, and web dashboard. Docker Compose, MIT. ![GitHub Repo stars](https://img.shields.io/github/stars/alamparelli/alf?style=social)
      • OpenAgent - Self-hostable personal assistant with LLM + RAG, loops for desktop/browser/coding, MCP and many providers. ![GitHub Repo stars](https://img.shields.io/github/stars/the-open-agent/openagent?style=social)
      • Everything OpenAI Codex - Open-source workflow system for OpenAI Codex that bundles agents, skills, commands, hooks, memory patterns, install profiles, and validation checks for repeatable coding sessions. ![GitHub Repo stars](https://img.shields.io/github/stars/mturac/everything-openai-codex?style=social)
      • Alfred - Self-hosted runtime for autonomous Claude Code and Codex agents that turns GitHub issues into reviewed pull requests. Per-firing git worktrees, label-driven state, role-based engine routing, Slack reports. Python, MIT, macOS/Linux. ![GitHub Repo stars](https://img.shields.io/github/stars/luminik-io/alfred-os?style=social)
      • career-ops - AI-powered job search orchestrator built on Claude Code. 14-skill pipeline that evaluates jobs, generates ATS-tailored PDFs, and tracks applications. Local-first, MIT. ![GitHub Repo stars](https://img.shields.io/github/stars/santifer/career-ops?style=social)
      • Hivekeep - Self-hosted platform of autonomous, persistent personal AI agents that collaborate, remember across months, and build their own tools. Multi-channel (Telegram, WhatsApp, Slack, Discord, Signal, Matrix), single container with Bun and SQLite. ![GitHub Repo stars](https://img.shields.io/github/stars/MarlBurroW/hivekeep?style=social)
      • AIDE - ML-engineering agent that uses tree search to optimize code against an eval metric, reaching human-level performance on Kaggle/MLE-bench. ![GitHub Repo stars](https://img.shields.io/github/stars/WecoAI/aideml?style=social)
      • Darkmoon - Open source autonomous AI penetration testing platform where Markdown methodology agents orchestrate 80+ offensive security tools through MCP controlled execution with agentic reasoning, keeping an evidence trail per finding. Model agnostic, tuned for Claude Opus. ![GitHub Repo stars](https://img.shields.io/github/stars/ASCIT31/Dark-Moon?style=social)
      • LoopTroop - Local GUI orchestrator for AI coding agents where an LLM Council plans, atomic beads execute in isolated git worktrees, and a Ralph Loop retries failures with fresh context. ![GitHub Repo stars](https://img.shields.io/github/stars/looptroop-ai/LoopTroop?style=social)
      • Tura - AGPL-3.0 local coding agent with CLI/TUI/GUI, task-scoped context, macro command execution, verification, and public benchmark artifacts. ![GitHub Repo stars](https://img.shields.io/github/stars/Tura-AI/tura?style=social)
      • Caesar - Autonomous research agent that builds a knowledge graph during web exploration via a Perceive-Think-Act loop, then refines drafts through adversarial artifact synthesis. Multi-provider via litellm. ![GitHub Repo stars](https://img.shields.io/github/stars/jasonzliang/caesar-agent?style=social)
      • CompozyOS - Self-hosted agent operating system: 26 providers, background loops and schedules, shared memory, approvals and an agent-to-agent network.
      • BitFun - Open-source coding agent with a Rust runtime, desktop and CLI interfaces, self-hosted multi-device control, and stateful Mini Apps. ![GitHub Repo stars](https://img.shields.io/github/stars/GCWing/BitFun?style=social)
      • Ouroboros - Self-hosted general-purpose agent with durable identity and memory, reviewed self-modification, specialist subagent swarms, and desktop or headless operation. ![GitHub Repo stars](https://img.shields.io/github/stars/razzant/ouroboros?style=social)
      • OpenDraft - Autonomous research-writing agent: 19 specialized agents turn a prompt into a long-form, source-grounded draft with citations verified against CrossRef, OpenAlex and Semantic Scholar. PDF/DOCX/LaTeX export, 57+ languages, bring-your-own model keys. ![GitHub Repo stars](https://img.shields.io/github/stars/federicodeponte/opendraft?style=social)
      • FutureOS - One AI agent everywhere you work: terminal UI, desktop, mobile, CLI, and IM bots from a single Rust backend, with approval-gated tools and a loop control plane for 24h+ runs. ![GitHub Repo stars](https://img.shields.io/github/stars/futuregene/future-os?style=social)
      • Sudarshan - Durable build harness that drives an LLM from an idea, PRD, or spec to software gated on passing verification commands, with resumable checkpointed state; provider-neutral across OpenAI-compatible, Anthropic, Gemini, local, and command-bridge backends. ![GitHub Repo stars](https://img.shields.io/github/stars/Suraj1235/sudarshan-superharness?style=social)
      • SARA - Self-hosted WhatsApp AI agent (AGPL-3.0) with 20 industry verticals, function calling (30+ tools), RAG via pgvector, and multi-provider LLM failover (Groq โ†’ Cerebras โ†’ SambaNova โ†’ Mistral). ![GitHub Repo stars](https://img.shields.io/github/stars/Alessandro114/sara?style=social)
      • Atomic Agent - Local-first CLI and TUI coding agent that runs open-weight models entirely on your machine via a llama.cpp fork. 56 tools (browser, filesystem, git, memory, vision), MCP support, and a 5-layer local memory. macOS/Linux/Windows, MIT. ![GitHub Repo stars](https://img.shields.io/github/stars/AtomicBot-ai/atomic-agent?style=social)
      • Agent Swarm - Self-hosted multi-agent system where a lead agent delegates tasks to specialized workers with shared memory, tools, schedules, and review gates. ![GitHub Repo stars](https://img.shields.io/github/stars/desplega-ai/agent-swarm?style=social)
      • Tracefold - Verified transformation calculus, pre-commit inverse escrow, and offline DSSE receipts for AI agent tool executions and filesystem mutations. ![GitHub Repo stars](https://img.shields.io/github/stars/TraceFold/tracefold?style=social)
      • Aster - Open-source terminal coding agent that reads your code, answers questions, edits files, runs commands, and reviews your changes. Works with any OpenAI-compatible provider (OpenRouter, OpenAI, Groq, Anthropic, local models). Rust, Apache-2.0. ![GitHub Repo stars](https://img.shields.io/github/stars/Zfinix/aster?style=social)
      • SDP (Social Daily Poster) - Self-hosted agent that harvests your Claude Code or Codex sessions, screenshots and messages each night, drafts one social post from the day's work, and routes it through a private Telegram bot for approval before it publishes to LinkedIn, X or Reddit. Python 3, MIT. ![GitHub Repo stars](https://img.shields.io/github/stars/dimamak/sdp?style=social)
      • ENZO - Self-hosted BYOK AI workspace with multi-provider chat (Groq, OpenRouter, NVIDIA, Hugging Face, Google AI), agents with scheduled runs, and skills like Gmail, Google Calendar, web search, and project generation. Provider keys are sealed client-side with AES-256-GCM and no middleman service is involved. Apache-2.0, Docker deployment. ![GitHub Repo stars](https://img.shields.io/github/stars/theguysudo/ENZO?style=social)
      • Kapso - Self-improving software factory for AI/ML objectives. State an objective and it runs a campaign: candidate solutions designed, implemented by coding agents (Claude Code, Codex), measured against the objective, and the closest refined until it is met. Each finished campaign leaves lessons carrying the evidence that earned them, and repositories and papers feed the same knowledge hub, so the next campaign starts from what earlier work established. ![GitHub Repo stars](https://img.shields.io/github/stars/Leeroo-AI/kapso?style=social)
      • PI-Desktop - Local-first desktop workspace for AI coding agents with persistent projects and sessions, model switching, Plan/Goal modes, permission controls, plugins, MCP, and multi-agent orchestration. ![GitHub Repo stars](https://img.shields.io/github/stars/vastsa/PI-Desktop?style=social)
    • Multi-Agent Task Solver Projects

      • MetaGPT - ๐ŸŒŸ The Multi-Agent Framework: Given one line Requirement, return PRD, Design, Tasks, Repo ![GitHub Repo stars](https://img.shields.io/github/stars/geekan/MetaGPT?style=social)
      • ChatDev - Create Customized Software using Natural Language Idea (through LLM-powered Multi-Agent Collaboration) ![GitHub Repo stars](https://img.shields.io/github/stars/OpenBMB/ChatDev?style=social)
      • DevOpsGPT - Multi agent system for AI-driven software development. ![GitHub Repo stars](https://img.shields.io/github/stars/kuafuai/DevOpsGPT?style=social)
      • Giselle - Giselle is an agentic workflow builder that empowers you to create AI-driven solutions with ease. ![GitHub Repo stars](https://img.shields.io/github/stars/giselles-ai/giselle?style=social)
      • EvoAgentX - EvoAgentX is building a Self-Evolving Ecosystem of AI Agents, it will give you automated framework for evaluating and evolving agentic workflows. ![GitHub Repo stars](https://img.shields.io/github/stars/EvoAgentX/EvoAgentX?style=social)
      • GenoMAS - Multi-agent framework for robust automation of scientific analysis workflows, such as gene expression analysis. ![GitHub Repo stars](https://img.shields.io/github/stars/Liu-Hy/GenoMAS?style=social)
      • RadOps - RadOps is an AI-powered, multi-agent platform that automates DevOps workflows with human-level reasoning. ![GitHub Repo stars](https://img.shields.io/github/stars/mehrdadrad/radops?style=social)
      • Hivemoot - Framework for AI agent teams that build real software on GitHub โ€” agents get roles, propose features, vote, review code, and ship autonomously. Colony is the first project built this way. ![GitHub Repo stars](https://img.shields.io/github/stars/hivemoot/hivemoot?style=social)
      • Orchard Kit - Six zero-dependency Python modules for autonomous agent governance and cognitive architecture: runtime security, confabulation detection, self-audit, agent discovery, cognitive architecture, and collective cognition. ![GitHub Repo stars](https://img.shields.io/github/stars/OrchardHarmonics/orchard-kit?style=social)
      • ClawFleet - Self-hosted AI fleet management with browser dashboard, Docker isolation, and bot-to-bot collaboration in Discord. ![GitHub Repo
      • Shire - Persistent workspaces for AI agent teams with inter-agent mailboxes, shared drive, and full context preservation. Supports Claude Code, OpenCode, Pi Agent and more. ![GitHub Repo stars](https://img.shields.io/github/stars/victor36max/shire?style=social)
      • Bernstein - Deterministic orchestrator that spawns parallel coding agents (Claude Code, Codex CLI, Gemini CLI), verifies with tests, and auto-commits. Zero LLM tokens on coordination. ![GitHub Repo stars](https://img.shields.io/github/stars/chernistry/bernstein?style=social)
      • Maestro Orchestrate - Multi-agent development orchestration platform coordinating 22 specialized AI agents through 4-phase workflows with native parallel execution, persistent sessions, and least-privilege security tiers across Gemini CLI, Claude Code, and Codex. ![GitHub Repo stars](https://img.shields.io/github/stars/josstei/maestro-orchestrate?style=social)
      • swarm-orchestrator - Contract-first multi-agent orchestrator that races persona candidates per typed obligation, verifies before commit, and logs every action in an append-only hash-chained ledger; deterministic offline default with optional Claude, Codex, Copilot, and local LLM (Ollama, llama.cpp, vLLM) providers. Ships with a GitHub Action. ![GitHub Repo stars](https://img.shields.io/github/stars/moonrunnerkc/swarm-orchestrator?style=social)
      • Maestro - Open-source desktop command center for running multiple AI coding agents (Claude Code, Codex, Gemini CLI, etc.) in parallel, with Cue event automation, Auto Run playbooks, Group Chat across local and remote agents, and a maestro-cli that agents can drive themselves. ![GitHub Repo stars](https://img.shields.io/github/stars/RunMaestro/Maestro?style=social)
      • OpenBusiness - Multi-agent pipeline that turns a company name + domain into an evidence-labeled business model report; runs JTBD, value proposition, GTM, unit economics, moat, canvas synthesis, and an assumption stress test, tagging every claim verified, inferred, or missing. Built on LangGraph; deterministic unit economics in Python. ![GitHub Repo stars](https://img.shields.io/github/stars/wanikua/OpenBusiness?style=social)
      • NextRole - A supervisor agent coordinates three sub-agents (hiring-recon, resume-tailor, interview-coach) to turn a CV and job description into a tailored resume, interview-prep doc, and day-of battlecard. ![GitHub Repo stars](https://img.shields.io/github/stars/tam159/next-role?style=social)
      • OpenAcme - Local-first AI workforce platform โ€” named agents with roles, personas, tools, memory, and per-agent MCP servers that self-organize through task delegation. Any agent can assign work to another; the scheduler wakes coworkers when dependencies clear. ![GitHub Repo stars](https://img.shields.io/github/stars/sandydasari/openacme?style=social)
      • h5i - CLI that runs several coding agents (Claude Code, Codex) on the same task, each in an isolated git worktree sandbox, has them peer-review each other, then a neutral verifier replays every candidate, runs the tests itself, and merges the one that passes. Run metadata is versioned in the repo under refs/h5i/*. Rust, Apache-2.0. ![GitHub Repo stars](https://img.shields.io/github/stars/h5i-dev/h5i?style=social)
      • AionUi - Open-source desktop client that runs multiple agent CLIs (Claude Code, Codex, Gemini CLI, Qwen Code) side by side, with multi-session chat, MCP and ACP support, and local file management. ![GitHub Repo stars](https://img.shields.io/github/stars/iOfficeAI/AionUi?style=social)
      • Orkas - MIT-licensed, local-first multi-agent desktop application where a Commander coordinates specialist agents for research, coding, data analysis, documents, and media. ![GitHub Repo stars](https://img.shields.io/github/stars/Orkas-AI/Orkas?style=social)
      • claude-consensus - Consensus protocol for LLM agents running on separate machines: propose/counter/accept/commit rounds, a dual-rail message bus with ACK tracking, and self-healing sync to keep agent state from drifting. ![GitHub Repo stars](https://img.shields.io/github/stars/tonydzi/claude-consensus?style=social)
      • Vicoa - Agentic IDE and AI orchestrator for running Claude Code, Codex, OpenCode, Gemini, Cursor, GitHub Copilot, Kimi, and Hermes agents in parallel, each in its own git worktree, steered from a unified dashboard with real-time mobile sync and push notifications. ![GitHub Repo stars](https://img.shields.io/github/stars/vicoa-ai/vicoa?style=social)
      • ClawFleet - Self-hosted AI fleet management with browser dashboard, Docker isolation, and bot-to-bot collaboration in Discord. ![GitHub Repo
      • Agent Teams - Open-source desktop app for coordinating multi-agent workflows with Kanban task management, agent messaging, code review, and approval controls. ![GitHub Repo stars](https://img.shields.io/github/stars/777genius/agent-teams-ai?style=social)
      • Bernstein - Open-source governance layer for AI agents. No model in the coordination loop: deterministic scheduling, per-task git worktree isolation, byte-identical replay, signed lineage, and an opt-in HMAC audit chain. Drives 40+ CLI coding agents (Claude Code, Codex CLI, Gemini CLI). Apache-2.0. ![GitHub Repo stars](https://img.shields.io/github/stars/sipyourdrink-ltd/bernstein?style=social)
      • Bunkhouse - Self-hosted multitenant platform for AI employees with company inbox, org chart, and governed procedures. ![GitHub Repo stars](https://img.shields.io/github/stars/braedonsaunders/bunkhouse?style=social)
      • XYZZY - Self-hosted multiplayer workspace where a team branches a question into parallel specialist agent runs, includes or excludes each output, and publishes a Decision Brief with every claim linked to its source output; hash-chained event log, one Python process on SQLite, works with any OpenAI-compatible endpoint. ![GitHub Repo stars](https://img.shields.io/github/stars/Project-Nexus-YR/XYZZY?style=social)
      • YYLO - Kanban-driven CLI orchestrator that runs coding agents (Claude Code, Codex, Gemini CLI) in parallel across isolated git worktrees, with a merge queue that reviews and merges verified task work. ![GitHub Repo stars](https://img.shields.io/github/stars/yylo-dev/yylo?style=social)
    • Tools

      • Metorial - Integration gateway that links AI agents to 600+ tools via unified MCP/OAuth interface with built-in scaling and monitoring. ![GitHub Repo stars](https://img.shields.io/github/stars/metorial/metorial?style=social)
      • musecl-memory - Zero-dependency file-based memory sync for AI agents using bash, git, and markdown. Lightweight alternative to vector DBs for agent persistence. ![GitHub Repo stars](https://img.shields.io/github/stars/musecl/musecl-memory?style=social)
      • APort Agent Guardrails - Pre-action authorization for OpenClaw and agent frameworks. `before_tool_call` plugin, 40+ blocked patterns, local or API. Setup: `npx @aporthq/agent-guardrails` ![GitHub Repo stars](https://img.shields.io/github/stars/aporthq/aport-agent-guardrails?style=social)
      • Agent OS - A kernel architecture for governing autonomous AI agents. Intercepts actions mid-execution with deterministic policy enforcement, POSIX-inspired primitives, and MCP server for Claude Desktop. ![GitHub Repo stars](https://img.shields.io/github/stars/imran-siddique/agent-os?style=social)
      • AgentGuard - Lightweight observability and runtime guardrails for AI agents โ€” loop detection, budget enforcement, cost tracking, and deterministic replay. Zero dependencies, LangChain integration. ![GitHub Repo stars](https://img.shields.io/github/stars/bmdhodl/agent47?style=social)
      • Agent OS - A kernel architecture for governing autonomous AI agents. Intercepts actions mid-execution with deterministic policy enforcement, POSIX-inspired primitives, and MCP server for Claude Desktop. ![GitHub Repo stars](https://img.shields.io/github/stars/microsoft/agent-governance-toolkit?style=social)
      • WFGY 16 Problem Map - Framework-agnostic debugging and evaluation checklist for LLM agents and RAG systems, with a practical 16-problem failure map covering retrieval, vector store, prompt / tool contract, and deployment issues in real workflows. ![GitHub Repo stars](https://img.shields.io/github/stars/onestardao/WFGY?style=social)
      • WritBase - MCP-native task management control plane for AI agent fleets with multi-agent permissions, delegation safety, and full provenance. ![GitHub Repo stars](https://img.shields.io/github/stars/Writbase/writbase?style=social)
      • Cortex - Persistent AI memory for coding assistants. Auto-captures decisions, patterns, and context across sessions. VSCode extension + CLI + MCP server. ![GitHub Repo stars](https://img.shields.io/github/stars/SKULLFIRE07/cortex-memory?style=social)
      • Steel Browser - Open-source browser infrastructure for AI agents and apps, supporting session-backed web automation, extraction, screenshots, and PDFs. ![GitHub Repo stars](https://img.shields.io/github/stars/steel-dev/steel-browser?style=social)
      • BGPT MCP - MCP server for searching scientific papers and retrieving structured experimental data extracted from full-text studies. ![GitHub Repo stars](https://img.shields.io/github/stars/connerlambden/bgpt-mcp?style=social)
      • Agent Brain - 7-layer cognitive memory system for AI agents with perception gate, dream cycle, and predictive capabilities. Self-hostable via Docker. ![GitHub Repo stars](https://img.shields.io/github/stars/kaderosio/agent-brain?style=social)
      • CueAPI - Open source scheduling and execution accountability API for AI agents. Retries, outcome tracking, and alerts when agents fail silently. ![GitHub Repo stars](https://img.shields.io/github/stars/cueapi/cueapi-core?style=social)
      • Uni-CLI - Universal CLI for AI agents โ€” 756 commands across 167 sites (web, desktop, Electron apps). Self-repairing 20-line YAML adapters, auto-JSON output, ~80 tokens per call. TypeScript, Apache-2.0. ![GitHub Repo stars](https://img.shields.io/github/stars/olo-dot-io/Uni-CLI?style=social)
      • clideck - WhatsApp-like dashboard for managing multiple AI coding agents in one browser window. Live status, session resume, autopilot that routes work between agents, and mobile remote. ![GitHub Repo stars](https://img.shields.io/github/stars/rustykuntz/clideck?style=social)
      • DexPaprika MCP - Open-source MCP server for querying decentralized exchange data across 34 blockchains. Exposes pool details, token metadata, OHLCV charts, trade history, and real-time swap streams via SSE. ![GitHub Repo stars](https://img.shields.io/github/stars/coinpaprika/dexpaprika-mcp?style=social)
      • Desktop Control - CLI tool for AI agents to control macOS apps via screen, mouse, and keyboard. GPU-accelerated, local OCR and vision, works with any AI model. ![GitHub Repo stars](https://img.shields.io/github/stars/yaroshevych/desktopctl?style=social)
      • KubeStellar Console - Multi-cluster Kubernetes dashboard with AI operations agent (kc-agent) that bridges LLMs to live clusters via MCP for AI-assisted troubleshooting, observability, and management across edge and cloud. CNCF Sandbox project. ![GitHub Repo stars](https://img.shields.io/github/stars/kubestellar/console?style=social)
      • BrowserTrace - Local flight recorder for AI browser agents with screenshots, URLs, model I/O, failure timelines, and public-safe HTML exports. ![GitHub Repo stars](https://img.shields.io/github/stars/aaronlab/browsertrace?style=social)
      • Kontext CLI - Open-source CLI for local guardrails, risk scoring, and redacted tool-call traces for AI agent sessions. ![GitHub Repo stars](https://img.shields.io/github/stars/kontext-security/kontext-cli?style=social)
      • agenttrace - Local-first TUI observability for AI coding agent sessions, with cost, token, tool failure, latency, anomaly, health score, diff, and CI gate views. ![GitHub Repo stars](https://img.shields.io/github/stars/luoyuctl/agenttrace?style=social)
      • AgentSkeptic - Verifies AI agent workflows by checking real database state instead of logs or traces. ![GitHub Repo stars](https://img.shields.io/github/stars/jwekavanagh/agentskeptic?style=social)
      • Agent-Wiz - Python CLI by Repello AI for extracting agentic workflows from LangChain/LangGraph/CrewAI/AutoGen and running automated threat modeling against the resulting graphs. ![GitHub Repo stars](https://img.shields.io/github/stars/Repello-AI/Agent-Wiz?style=social)
      • MisakaNet - Git-based shared memory for AI agents. Cross-agent lesson/knowledge sync via GitHub Issues. "Lessons learned. Lessons shared." ![GitHub Repo stars](https://img.shields.io/github/stars/Ikalus1988/MisakaNet?style=social)
      • authsome - Local credential broker for AI agents. Log in once via OAuth2 or API key, vault stores secrets locally, local proxy injects them at request time so agents never see the raw values. 45 providers bundled. ![GitHub Repo stars](https://img.shields.io/github/stars/agentrhq/authsome?style=social)
      • Perseus - Live workspace context engine for AI agents. Renders AGENTS.md at session start. Plug-in for Claude Code, Codex, Hermes.
      • codex-profiles - Bash CLI for switching OpenAI Codex CLI/Desktop accounts with isolated `CODEX_HOME` profiles. ![GitHub Repo stars](https://img.shields.io/github/stars/Ducksss/codex-profiles?style=social)
      • CommonGround Kernel - PostgreSQL-backed shared work substrate for human-agent and multi-agent systems, with durable public work records, handoff facts, causal lineage, claim fencing, and pull-first recovery across runtimes. ![GitHub Repo stars](https://img.shields.io/github/stars/Intelligent-Internet/CommonGround?style=social)
      • WinkTerm - Self-hosted AI terminal where the agent shares your PTY session; in-terminal `#` chat, SSH, and HTTP Agent API with installable skill for coding agents. ![GitHub Repo stars](https://img.shields.io/github/stars/Cznorth/winkterm?style=social)
      • DOS (dos-kernel) - Trust kernel for AI agent fleets: verifies an agent's "done" claim from git evidence (never self-report), arbitrates file collisions between concurrent agents, and refuses with structured machine-checkable reasons. CLI + MCP server + Claude Code plugin. ![GitHub Repo stars](https://img.shields.io/github/stars/anthony-chaudhary/dos-kernel?style=social)
      • EGC - Cross-session persistent memory layer for AI coding agents (Claude Code, Cursor, Gemini CLI, Codex, Windsurf, Amp, Kiro, and more). SQLite-backed. ![GitHub Repo stars](https://img.shields.io/github/stars/Fmarzochi/EGC?style=social)
      • ax - Local telemetry for AI coding agents.
      • Cynative - Agentic security CLI that runs code in a built-in sandbox to research cloud, code and runtime. Read-only by construction. ![GitHub Repo stars](https://img.shields.io/github/stars/cynative/cynative?style=social)
      • Tree Ring Memory - Local-first Rust CLI and TUI for AI agent memory lifecycle with SQLite/FTS recall, audit, consolidation, forgetting, and framework discovery. ![GitHub Repo stars](https://img.shields.io/github/stars/TerminallyLazy/Tree-Ring-Memory?style=social)
      • poolsplit - Pool-split retrieval for agent memory: reserved per-type token budgets so low-priority entries surface and behavioral corrections never get crowded out. Zero dependencies, pluggable scorer. ![GitHub Repo stars](https://img.shields.io/github/stars/SpicyNoodles3/poolsplit?style=social)
      • Nika - Intent-as-code workflow engine for AI agents: reviewable YAML DAGs statically checked (schema, permits, honest cost floor) before any token is spent, tamper-evident traces after. Single Rust binary. ![GitHub Repo stars](https://img.shields.io/github/stars/supernovae-st/nika?style=social)
      • Caspian - One messaging identity for an AI agent across Slack, Discord, Telegram, Instagram, email, and X โ€” a single `on_message` handler with threading, webhook verification, and platform quirks handled. Python + TypeScript SDK. ![GitHub Repo stars](https://img.shields.io/github/stars/TryCaspian/caspian-sdk?style=social)
      • DSH Studio - Cross-platform desktop host for installing, running, health-checking, and supervising DeepSeek Harness locally. ![GitHub Repo stars](https://img.shields.io/github/stars/Moresyl/dsh-studio?style=social)
      • Lians - Local-first memory layer for AI agents with MCP, Python, and TypeScript interfaces; SQLite-backed recall, user-controlled inspection/correction/deletion, and point-in-time memory receipts. ![GitHub Repo stars](https://img.shields.io/github/stars/Lians-ai/Lians?style=social)
      • Hexis - Git-backed platform for managing and sharing skills, tools, and context across AI agents through a remote MCP server. ![GitHub Repo stars](https://img.shields.io/github/stars/Bevel-Software/Hexis?style=social)
      • Caura - Governed shared memory for AI agent fleets, with multi-agent and multi-tenant support, MCP integration, trust tiers, audit trails, knowledge graph capabilities, and self-improving retrieval. ![GitHub Repo stars](https://img.shields.io/github/stars/caura-ai/caura?style=social)
      • Statewave - Open-source memory runtime for AI agents, providing durable, structured, provenance-tagged context with deterministic, token-bounded memory retrieval. ![GitHub Repo stars](https://img.shields.io/github/stars/smaramwbc/statewave?style=social)
      • Open Index - Structured context layer for domain-specific agents with typed knowledge graphs, hybrid search, and read/write MCP access. ![GitHub Repo stars](https://img.shields.io/github/stars/DrDroidLab/open-index?style=social)
      • Compartment - Local-first, offline encrypted vector memory for AI agents over MCP or CLI, with AEAD-encrypted-at-rest records and embeddings, RAM-resident exact vector search, per-record crypto-shred deletion, and a hash-chained audit log. Apache-2.0, Python. ![GitHub Repo stars](https://img.shields.io/github/stars/MaxFreedomPollard/Compartment?style=social)
      • MCP Lens - Open-source DeepSeek Harness plugin that discovers MCP tools through search and invokes selected tools with their exact input schemas. ![GitHub Repo stars](https://img.shields.io/github/stars/labmimors/dsh-mcp-lens?style=social)
      • Agent Coordinator - Codex skill that records complex tasks as revisioned work graphs, rejects overlapping write scopes, reconciles uncertain work before retry, and reruns completion checks. ![GitHub Repo stars](https://img.shields.io/github/stars/alanhoff/agent-coordinator?style=social)
      • Oathra - Apache-2.0 TypeScript runtime for phone agents with a standalone evidence-verification engine; anchors result fields to callee utterances and applies deterministic completion rules, with a simulator and adversarial tests. ![GitHub Repo stars](https://img.shields.io/github/stars/FORIFOR/oathra?style=social)
      • OrcaReplay - Records a coding agent below the harness โ€” model traffic, shell exit codes, per-turn file changes and MCP calls on one timeline โ€” then replays the run offline with the network off, or forks it from any checkpoint onto a different model.
      • ValetFS - Secrets stay on a paired device and are lent to the agent's machine into daemon memory only, served over FUSE or loopback WebDAV; the daemon zero-wipes them when the pairing drops or a grace window expires. ![GitHub Repo stars](https://img.shields.io/github/stars/winm2m/valet-fs?style=social)
      • sofagent - Audit-first governance harness for AI coding agents: 24 rules enforced at commit time via git hooks, HMAC-chained audit log, snapshot rollback. MIT. ![GitHub Repo stars](https://img.shields.io/github/stars/KongFangXun/sofagent?style=social)
      • Webcmd - Self-learning browser infrastructure for AI agents: learns a site's navigation once, then compiles it into deterministic per-site CLI commands. TypeScript, Apache-2.0. ![GitHub Repo stars](https://img.shields.io/github/stars/agentrhq/webcmd?style=social)
      • 5dive - Self-hosted CLI that runs a team of coding agents on one Linux host: each agent is its own Linux user running `claude`, `codex`, `opencode`, `hermes` or another CLI as a systemd service, coordinating through an org chart and a shared SQLite backlog, escalating to Telegram only when a human must decide. MIT. ![GitHub Repo stars](https://img.shields.io/github/stars/5dive-ai/5dive?style=social)
      • Ordewell - Open-source terminal CLI and TUI that turns one goal into an ordered, editable plan of coding-agent tasks, each with its own runner (Claude Code, Codex, OpenCode), model and mode; a task is complete only when a completion marker appears in the runner's output. Apache-2.0. ![GitHub Repo stars](https://img.shields.io/github/stars/ordewell/ordewell?style=social)
      • Busabase - Open-source database and workspace for AI agents to manage typed tables, fields, views, records, docs, files, and search; writes can become ChangeRequests for human review. Streamable HTTP MCP server, local-first with PGlite, and self-hostable. MIT. ![GitHub Repo stars](https://img.shields.io/github/stars/busabase/busabase?style=social)
  • Benchmark/Evaluator

    • Advanced Components

      • AgentBench - A Comprehensive Benchmark to Evaluate LLMs as Agents ![GitHub Repo stars](https://img.shields.io/github/stars/THUDM/AgentBench?style=social)
      • agentops - Python SDK for agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks like CrewAI, Langchain, and Autogen ![GitHub Repo stars](https://img.shields.io/github/stars/AgentOps-AI/agentops?style=social)
      • langtrace - Langtrace ๐Ÿ” is an open-source, Open Telemetry based end-to-end observability tool for LLM applications, providing real-time tracing, evaluations and metrics for popular LLMs, LLM frameworks, vectorDBs and more.. Integrate using Typescript, Python. ![GitHub Repo stars](https://img.shields.io/github/stars/Scale3-Labs/langtrace?style=social)
      • ToolBench - An open platform for training, serving, and evaluating large language model for tool learning. ![GitHub Repo stars](https://img.shields.io/github/stars/OpenBMB/ToolBench?style=social)
      • LLM-Agent-Benchmark-List - A benchmark list for evaluation of large language models.
      • open-operator-evals
      • GenoTEX - A benchmark for evaluating LLM agents on end-to-end gene expression data analysis, featuring comprehensive gene-trait association analysis with expert-curated annotations. ![GitHub Repo stars](https://img.shields.io/github/stars/Liu-Hy/GenoTEX?style=social)
    • Tools

      • LiveMCP-101 - Benchmark of 101 real-world MCP tool-use queries with plan-based evaluation highlighting agent orchestration gaps.
      • AgentLab - Open-source framework for developing and evaluating web agents with benchmark-driven workflows. ![GitHub Repo stars](https://img.shields.io/github/stars/ServiceNow/AgentLab?style=social)
      • BrowserGym - Gym-style benchmark and environment toolkit for evaluating browser-using web agents. ![GitHub Repo stars](https://img.shields.io/github/stars/ServiceNow/BrowserGym?style=social)
      • OSWorld - Benchmark for multimodal desktop computer-use agents with tasks across real operating-system environments. ![GitHub Repo stars](https://img.shields.io/github/stars/xlang-ai/OSWorld?style=social)
      • ClawBench - Browser-agent benchmark of 153 everyday tasks on 144 live production websites across 15 categories; a submission-interception layer blocks the final write request for safe evaluation on real sites. ![GitHub Repo stars](https://img.shields.io/github/stars/reacher-z/ClawBench?style=social)
      • Cross-Agent Review Queue 2026 - Open dataset of cross-agent collaboration review transcripts (Codex <-> Claude reviewer / architect / implementer handoffs) with structured fields for owner-goal restatement, review lens, and result code (NEW_SIGNAL / NO_NEW_SIGNAL); useful for multi-agent handoff and review-quality evaluation.
      • Future AGI - Open-source platform to simulate, evaluate, trace, guardrail, and optimize LLM and AI agent apps, with 70+ eval metrics and OpenTelemetry-native tracing across 50+ frameworks. ![GitHub Repo stars](https://img.shields.io/github/stars/future-agi/future-agi?style=social)
      • Multi-SWE-bench - Multi-language extension of SWE-bench for evaluating software engineering agents beyond Python repositories. ![GitHub Repo stars](https://img.shields.io/github/stars/multi-swe-bench/multi-swe-bench?style=social)
      • CIAgent - Pytest-native regression testing for AI agents โ€” golden-trace diffing, cost guardrails, multi-run stability scoring with flip attribution, LLM-judge auditing, and one-command import of production traces (OTel/Langfuse/LangSmith) into CI tests. ![GitHub Repo stars](https://img.shields.io/github/stars/suniel12/ciagent?style=social)
      • Sabot - Injects one controlled fault into a running LangGraph, CrewAI or AutoGen/Magentic-One pipeline โ€” corrupted tool result, falsified success report, altered inter-agent message, silent model downgrade, stale context, silent no-op โ€” and scores whether the pipeline's own reviewer, guardrail and orchestrator surfaces detect it. Pre-registered spec and adjudication anchors, deterministic scoring with no LLM in the headline path, full raw trace corpus published. ![GitHub Repo stars](https://img.shields.io/github/stars/Jott2121/sabot?style=social)
      • whatbroke - CLI that diffs two runs of an AI agent to show changes in tool calls, arguments, cost, latency, and outcomes, with multi-sample flake detection to demote pre-existing flakiness. ![GitHub Repo stars](https://img.shields.io/github/stars/arthi-arumugam-git/whatbroke?style=social)
      • ClawBench - Browser-agent benchmark of 281 everyday tasks (V1 152 + V2 129) on 163 live production websites across 15 categories; two-stage scoring โ€” a submission-interception layer blocks the final write request for safe evaluation on real sites, then an LLM judge checks the captured payload against the instruction. ![GitHub Repo stars](https://img.shields.io/github/stars/TIGER-AI-Lab/ClawBench?style=social)
      • AgentLeak - Python toolkit for evaluating privacy leakage across agent traces, including tool calls, inter-agent messages, shared memory, and logs, with redacted reports and CI gates. ![GitHub Repo stars](https://img.shields.io/github/stars/yagobski/agentleak?style=social)
      • SWE-bench - Benchmark for evaluating LLM systems on real-world GitHub issue resolution tasks. ![GitHub Repo stars](https://img.shields.io/github/stars/Princeton-NLP/SWE-bench?style=social)
      • agbenchmark - by AutoGPT
  • Frameworks

    • Advanced Components

      • langchain - โšก Building applications with LLMs through composability โšก ![GitHub Repo stars](https://img.shields.io/github/stars/langchain-ai/langchain?style=social)
      • awesome-langchain - ๐Ÿ˜Ž Awesome list of tools and projects with the awesome LangChain framework ![GitHub Repo stars](https://img.shields.io/github/stars/kyrolabs/awesome-langchain?style=social)
      • llama_index - LlamaIndex (formerly GPT Index) is a data framework for your LLM applications ![GitHub Repo stars](https://img.shields.io/github/stars/run-llama/llama_index?style=social)
      • agents - An Open-source Framework for Autonomous Language Agents ![GitHub Repo stars](https://img.shields.io/github/stars/aiwaves-cn/agents?style=social)
      • AutoGen - AutoGen is a framework that enables the development of LLM applications using multiple agents that can converse with each other to solve tasks. ![GitHub Repo stars](https://img.shields.io/github/stars/microsoft/autogen?style=social)
      • TaskWeaver - A code-first agent framework for seamlessly planning and executing data analytics tasks. ![GitHub Repo stars](https://img.shields.io/github/stars/microsoft/TaskWeaver?style=social)
      • AgentVerse - AgentVerse is designed to facilitate the deployment of multiple LLM-based agents in various applications. AgentVerse primarily provides two frameworks: task-solving and simulation. ![GitHub Repo stars](https://img.shields.io/github/stars/OpenBMB/AgentVerse?style=social)
      • SuperAGI - A dev-first open source autonomous AI agent framework. Enabling developers to build, manage & run useful autonomous agents quickly and reliably. ![GitHub Repo stars](https://img.shields.io/github/stars/TransformerOptimus/SuperAGI?style=social)
      • AutoChain - Build lightweight, extensible, and testable LLM Agents ![GitHub Repo stars](https://img.shields.io/github/stars/Forethought-Technologies/AutoChain?style=social)
      • modelscope-agent - An agent framework connecting models in ModelScope with the world ![GitHub Repo stars](https://img.shields.io/github/stars/modelscope/modelscope-agent?style=social)
      • Voice Lab - A comprehensive testing and evaluation framework for voice agents across language models, prompts, and agent personas. ![GitHub Repo stars](https://img.shields.io/github/stars/saharmor/voice-lab?style=social)
      • AgentSquare - Automatic LLM Agent Search In Modular Design Space ![GitHub Repo stars](https://img.shields.io/github/stars/tsinghua-fib-lab/AgentSquare?style=social)
      • MixedVoices - An Open source tool for analyzing and evaluating AI Voice agents. Track and visualize performance through call analysis and flow charts. Run complex simulations before pushing to production. ![GitHub Repo stars](https://img.shields.io/github/stars/mixedvoices/mixedvoices?style=social)
      • KaibanJS - KaibanJS is a JavaScript-native framework for building and managing multi-agent systems with a Kanban-inspired approach. ![GitHub Repo stars](https://img.shields.io/github/stars/kaiban-ai/KaibanJS?style=social)
      • Upsonic - Upsonic is a reliable agent framework supporting MCP, offering trusted agent workflows with verification layers. ![GitHub Repo stars](https://img.shields.io/github/stars/upsonic/upsonic?style=social)
      • notte - ๐Ÿ”ฅ Reliable Browser AI agents framework for building and deploying web automation agents with hybrid workflows combining AI and traditional scripting ![GitHub Repo stars](https://img.shields.io/github/stars/nottelabs/notte?style=social)
      • MixedVoices - An Open source tool for analyzing and evaluating AI Voice agents. Track and visualize performance through call analysis and flow charts. Run complex simulations before pushing to production. ![GitHub Repo stars](https://img.shields.io/github/stars/mixedvoices/mixedvoices?style=social)
      • Mastra - Mastra is an opinionated TypeScript framework that helps you build AI applications and features quickly. ![GitHub Repo stars](https://img.shields.io/github/stars/mastra-ai/mastra?style=social)
      • AppAgent - A novel LLM-based multimodal agent framework designed to operate smartphone applications. ![GitHub Repo stars](https://img.shields.io/github/stars/mnotgod96/AppAgent?style=social)
      • modelscope-agent - An agent framework connecting models in ModelScope with the world ![GitHub Repo stars](https://img.shields.io/github/stars/modelscope/modelscope-agent?style=social)
      • Swarms - Enterprise-grade multi-agent framework for orchestrating intelligent AI agents at scale. Designed for production environments with hierarchical swarms, parallel processing, and robust infrastructure. ![GitHub Repo stars](https://img.shields.io/github/stars/kyegomez/swarms?style=social)
      • LLMling-Agent - Multi-agent workflows and complex Agent interactions, both via YAML manifest and programmatic usage. Pydantic-AI and LiteLLM backends with human-in-the-loop integration ![GitHub Repo stars](https://img.shields.io/github/stars/phil65/llmling-agent?style=social)
      • superagent - ๐Ÿฅท The open framework for building AI Assistants ![GitHub Repo stars](https://img.shields.io/github/stars/homanp/superagent?style=social)