Category
Code agents
38,402 Code AI agents indexed on MeshKore — the most complete public catalog, ranked by popularity and updated daily.
38,402 agents · ranked by popularity · refine in the directory →
Code agents — page 316 of 385
一个 7×24 小时不间断运行,帮你跑实验、用 Git 维护公开仓库、更新 LaTeX(Overleaf)论文的科研自动化工作流。实测效率提升 25 倍。Claude Code Skill,一条命令装进任意项目。
A 股多智能体投研编排系统:7 个专职 subagent 以文件契约协作,长文档上下文隔离,数字 claim 双源核验,公司与行业双研究链路
A threat hunter that lives in your terminal with memories in Notion.
An AI-powered EEG classifier that detects emotional states during music listening and curates Spotify playlists grounded in what your brain actually did.
Claude Code / Codex skill for authoring and reviewing SuperDialog voice-agent playbook YAMLs — schema reference, 15-point failure-pattern checklist, interrupt-guard code pattern
Zero-dependency LLM agent runtime — tool use, MCP, sandboxed self-written tools, checkpoint/resume, approval gates, bi-temporal fact memory, code-level guards (context bounds, loop guard, egress control), declarative tool policy, OpenTelemetry. Not a framework. 755 tests, CI Linux/Windows, 22 runnable demos, online visual builder.
企业级 RAG 知识库问答系统 + AIOps 运维诊断 Agent(Plan-Execute-Replan / LangGraph / MCP)
Deterministic, inspectable workflows for coding agents. Kilin is a local, CLI-first workflow runtime: describe a workflow in plain files, validate it before anything runs, then execute it through Codex, Claude Code, or OpenCode with explicit approvals and persisted runs.
LibreControl — one universal /remote-control for every coding agent you run (Claude Code, Codex, Gemini, Cursor, Copilot, any MCP tool), across all your machines, centralized into the single chat app you choose: Telegram, WhatsApp, Slack, Signal, Discord. Self-hosted, with approvals and audit.
Engineering discipline for AI harnesses. Drop-in skill library for LLM coding agents.
Self-hosted, governed-autonomy SRE platform: an LLM agent that autonomously remediates infrastructure incidents behind a fail-closed prediction gate, mechanical verdicts, and a tamper-evident ledger. The agent can't act on a belief it hasn't checked.
Best AI Code Architect 2026: Diagram-First Agent Orchestrator with AST Editing & Atomic Commits
Give your AI agent a phone — a thin Python CLI over the BlueBubbles REST API to read, search, send, react, and attach in iMessage, with draft-then-confirm send gating and an agent-facing SKILL.md.
Benchmark code and interactive results for autonomous network-attack detection by tool-using LLM agents investigating raw PCAPs.
AI agents built spaCy model recommenders, then a fresh AI peer-review panel rejected all three
Post-response grounding check + cross-family adversarial review for LLM chat agents
Локальный модуль машинного зрения для LLM-агентов на чистом Rust: метрический 3D scene-graph каждый кадр, MCP-сервер, $0/CPU/приватно
Agent-ready project documentation template with START/STATE/TODO/DONE, workflows, lightweight tooling, and LLM-friendly project memory.
Benchmarked Claude Code plugin that makes cheaper Claude models (Haiku, Sonnet, Opus) match Fable 5's output at a fraction of the token cost — auto-routed reasoning levels, verification discipline, closed-loop habit correction, statusline HUD with usage meters, and a portable AGENTS.md export for any AI coding agent.
Research-first, brand-agnostic blog drafting engine for Claude Code — keyword research with SERP-ownership choice, an enforced adversarial editor, validated citations, and CMS-ready output.
Claude Code Workflow that executes cys:plan implementation plans with independent tasks running in parallel via a dependency DAG — ships the cys plugin (design, plan, run, check, ship).
An AI agent that falls deeper into any topic — books, papers, podcasts, and videos, all connected.
Runtime-neutral protocol for growing adaptive marketing agents: role + playbook + GEB learning, with schema validators and CI gates enforcing it. Pure spec — bring your own runtime (Codex, Claude Code, MCP, etc.).
LEASH-8: an 8-domain control model for AI agents with delegated authority. Scorecard, approval-design checklist, plan-vs-authorize pattern. Patterns from a live production agent operation, sanitized. Free, MIT.
Pediatric dose-extrapolation agent: a defensible starting dose for a child from adult pharmacokinetics via allometry × organ maturation, with a cited, graded, auditable rationale. Multi-agent (Opus + Sonnet, live PubMed/openFDA). Decision support, not prescribing.
Interactive AI coding agent using free OpenRouter models with real tool calling
Local AI Linux troubleshooting agent with llama.cpp, safe command execution, system diagnostics, Docker support, CLI, and web GUI.
163 specialist Claude Code agents + 24 routing skills + 16 strategy playbooks. Korean/Japanese Business Navigators, game-dev (Unity/Unreal/Godot/Roblox/Blender), XR/spatial, paid-media routing. skill-routing-arbitrator disambiguates the ~500-skill ecosystem. Fork of msitarzewski/agency-agents.
Evidence router & policy engine for coding agents — enforces source choice and proof-of-use through Claude Code hooks, including routes to MCP tools. Local-first; no LLM in the hook loop.
TeleAgent AI is an enterprise-grade, agentic omnichannel consumer intelligence engine engineered specifically for Deutsche Telekom AG (Germany & Europe) under Problem Statement 5: Omnichannel Consumer AI Engine for Digital Commerce.
GSD lifecycle orchestration as a native OpenClaw plugin — enforced research→plan→execute→verify→ship, hybrid retrieval, 0 Discord slots
Multi-agent collaboration engine with persistent, searchable, self-compacting memory. Agents research, critique and adjudicate over a shared SQLite memory with hybrid BM25 + embedding recall. Zero-dependency core, runs locally on Ollama.
Survivability-first quantitative research system. An AI council debates every architecture decision before code; deterministic, tested strategies do the trading. Walk-forward + purged CV + deflated Sharpe. LLMs never place trades.
Local-first Python toolkit for parallel Claude and LLM agent orchestration: consensus voting, stigmergy, boss-worker swarms, and benchmarks
AI-SOC — detects, explains and contains cyberattacks on critical infrastructure. Isolation Forest detection + RAG/MITRE attribution + policy-gated response, wired end-to-end. ET AI Hackathon 2026, PS7.
Let AI coding agents (Claude, Codex, Pi) see and control your Bevy game - remote input injection, screenshots, and frame-synced feedback over a simple JSON-lines TCP protocol.
Rich Brain / Clean Hands: a two-brain delegation pattern for LLM agents that keeps the deciding brain's context clean and exiles heavy work to an isolated executor.
More human oversight can make an AI agent less safe. Headroom is a human-in-the-loop firewall for coding agents that measures when to trust the human: oversight as resource allocation, not just classification.
Provider-neutral LLM router and model orchestration engine with auto-learning and budget zones
An agentic buyer's agent for an unrepresented Bay Area home purchase — deterministic Python core, LLM at the edges.
MCP server for Cognigy.AI - 132 tools that let Claude, Cursor & other AI assistants build, configure, test & operate conversational AI agents via the Model Context Protocol.
Use AI coding agents with LTspice via sim-cli: run circuits, inspect waveforms/logs, and produce replayable simulation artifacts.
Security & runtime checklist for moving LLM agents from demo to production: HITL gates, idempotency, prompt injection boundaries, tool risk classification (CC BY 4.0).
学习型项目:Agent Harness 与 RL 后训练学习路线(资料/笔记/动手锚点),含 learn-claude-code 主线,.claude 配置参考 RQ-TPP
Decision governance for code and the agents editing it — track PRD→ADR→SPEC decisions, map them to files, and gate changes against them. Deterministic, honest CLI + MCP.
Deterministic execution harness for AI agents — goals in, structured tasks out, verified results back.
Infer least-privilege MCP tool policies from agent traces
Graph RAG on SQLite for AI agents: vector retrieval + hand-curated wikilink graph + cross-encoder rerank, with a zero-token per-turn memory ledger. Working pilot.
Stop self-replicating prompt payloads from writing themselves into your AI agent's config files
ETH Zurich thesis project: a nine-agent AI pipeline for quant research, from hypothesis to backtest report
Judgment continuity for AI agents — load your project's scars, taste, and non-negotiables into every session. Judgment-layer companion to edda. (formerly project-doctrine)
AI STEM animation skill — ManimGL for coding agents
Financial Intelligence Workstation — A native, cross‑platform desktop terminal with 37 AI agents, 100+ data connectors, 16 broker integrations, QuantLib suite, visual node editor, and embedded Python analytics. Free and open‑source (MIT). Your thinking is the only limit.
AI-powered code review and autonomous development platform built with TypeScript, React, FastAPI, PostgreSQL and LLM-based agents.
A recipe for any AI agent to build a self-improving theory-of-mind model of its user from interaction logs (feedback precognition). What you mind is what you get.
Commitment tracking for LLM agents — a normative overlay that gives long-running agents a scoreboard of their own commitments, not just a memory of what happened.
Privacy-first AI assistant that runs entirely on your machine. Local Ollama inference, embedded SQLite/Kuzu/ChromaDB, Rust-backed Brain Firewall, MCP connectors.
Production 5-agent pipeline automating pre-sales: discovery → scoping → pricing → proposal. 6 BUs, 22 services. Groq + Llama 3.3. Adversarial CriticAgent.
Reproducible benchmark comparing fallback strategies for multi-agent LLM systems. Submitted to NeurIPS 2026 Workshop "Who Verifies the Agents?" (under review).
基于QQ的ai agent聊天服务
Make any coding agent work like a frontier model. Drop-in Agent Skills for disciplined planning, evidence-first debugging, and live-system safety — plus a Python project scaffold with an agent contract, review checklist, and definition of done. Model-agnostic, zero dependencies.
AI-powered content generation system using multi-agent collaboration. Scale your content production 10x with AI. Automated research, writing, and optimization for blogs and reports. From idea to publication-ready content in under 2 minutes. Open-source, self-hosted, production-grade.
Five ready-to-run CrewAI multi-agent crews — content team, market research, code documentation, SEO audit, product launch — on free NVIDIA NIM
Upstream pre-call budget gateway and tamper-evident SHA-256 audit ledger for autonomous AI agents (CrewAI, LangChain, OpenAI).
Multi agent AI debate simulation system built with CrewAI, FastAPI, and Next.js. Agents debate real-world topics using live web grounding and are evaluated by a structured AI judge panel, visualized through a strategic card based UI.
🤖 The insider's guide to AI agents that actually work. Staff picks, honest reviews, code examples, and learning paths. Enhanced from awesome-ai-agents.
The reconciling memory layer for AI agents (MCP server): a new value supersedes the old, "we dropped X" actually retracts X, and history stays queryable — so agents get current facts, not stale ones.
Multi-agent chain-of-thought system that turns a merchant pitch into a ready-to-publish promotional campaign.
Build autonomous AI agents that complete real-world tasks end-to-end - research, coding, data analysis, and finance agents.
Autonomous AI-powered vulnerability assessment platform — multi-agent bug hunting with CrewAI, real security tools, and real-time streaming
AutoGen integration for Scavio Search API -- real-time Google, Amazon, Walmart, YouTube, Reddit, TikTok, and Instagram search tools for AI agents
Real-world evaluation framework for AI coding agents — measures safety, containment, cost, and autonomy beyond just correctness.
🤖 Predict programming problem difficulty with AI using text analysis and machine learning for accurate complexity scoring.
An open-source autonomous QA agent that spawns AI-powered user agents and unleashes them on your product - web, mobile, or desktop. Each agent navigates independently, makes decisions, hits dead ends, finds bugs, and reports back. All without writing a single test script.
Deploy the Nous Research Hermes Agent (self-improving AI agent) to any cloud — Terraform modules for Hetzner, Oracle Always Free, GCP, AWS Lightsail, and DigitalOcean. One docker compose, five clouds, MIT licensed.
Free AI coding assistant for VS Code & Cursor — Claude, GPT-4, Gemini & 300+ models. No subscription.
AI-powered WiFi security tool with reinforcement learning, gene evolution, and local LLM on Raspberry Pi Zero 2W. Captures WPA/WPA2 handshakes autonomously with e-paper display. Inspired by Pwnagotchi + Bjorn + PicoClaw.
Autonomous, evidence-grounded web-pentest orchestrator. A LangGraph master dispatches adversarially-verified sub-agents that drive all traffic through the burpwn sandbox — deep, cross-chain, 0-false-positive.
Autonomous AI agent that bids on and completes coding jobs on the NEAR AI marketplace to earn NEAR tokens - engineering case study
Self-hosted autonomous coding agent with executable skills and memory that survives sessions
Self-hosted agent infrastructure framework — orchestration, hybrid RAG, workflows, tools and full-trace audit behind one API. Ships with a React console, SSE streaming, and support for any OpenAI-compatible LLM.
Give AI agents hands without handing them the keys: exact one-use approvals and verifiable receipts for registered actions.
Production-ready multi-agent orchestrator — confidence-gated routing with a pluggable classifier, MCP or in-process agent dispatch, per-agent circuit breakers, and pluggable session/breaker persistence (Firestore, Postgres, Redis).
A reasoning-first idle engine for persistent LLM agents. Behaviour emerges from the agent reasoning about its idle state, gated by a cheap model-free safety layer.
Personal Ai Agent
SoulMap AI: a content-first reflective companion with a curated Markdown knowledge base, Python detectors, and tooling to validate and bundle agent-ready skills.
Agent Discipline Layer (ADL) — keep a coding agent honest and make it prove its work: layered guardrails + a /goal contract Warden verifies against reality and signs. Claude Code today; Grok/Codex CLI/Pi planned. (formerly Claude Layers)
Async multi-agent AI library for Python — self-improving agent loops, APEX synthesis, bounded agent spawning, multimodal input, and true token streaming across 100+ LLM providers.
265 autonomous AI agents across 15 business domains. Full-stack SaaS framework with real model routing, RAG pipeline, guardrails, and Chrome extension. FastAPI + Next.js + K3s. Runs on any server. $0/month software cost.
Product Build Gate (hplan) + 5 agent-PM lifecycle plugins — 6 plugins, 43 skills, 18 commands for PMs who decide what, why, and how to build AI agents
Production-ready agents, your way.
Core engine for the autonomous development factory platform with workflow orchestration
TypeScript framework for orchestrating collaborative AI agents with task scheduling, shared memory, tools, and multi-provider LLM support.
An open framework for a persistent, self-regulating AI operations system on a coding-agent CLI: tiered memory + SSOT, security-first gates, a knowledge pipeline, multi-agent protocols, session survival, and a zero-dependency cockpit.
CodegniPy is a groundbreaking Python library that elevates AI to a first-class citizen of the language. It introduces a cognitive computing engine where deterministic code and non‑deterministic LLM reasoning coexist seamlessly, enabling you to write Python with intent rather than just instructions.
Multi-agent coding CLI with loop engineering — plan, act, verify, self-critique. Bridge Claude/Codex, MCP, skills, remote UI
One command to turn a developer into 100. 4-stage pipeline, NASA P10 quality gates, budget enforcement, and interactive dashboards. Built on Claude Code + Karpathy's context engineering.
Perpetual, self-correcting development loop for Claude Code: an RSI backlog loop + MAKER-style per-step error correction (decomposition + multi-agent voting, with your tests as the judge).
Portable Claude Code skills + agent panel that translate any project or files into any language — codebase i18n, WordPress/gettext, docs — with adversarial QA, per-run domain specialization, and terminology research.
Easy, fast, isolated microVM sandboxes for AI agents and untrusted code on AWS Lambda MicroVMs — Python SDK, asb CLI, and MCP server.