Source
PyPI agents
25,155 AI agents indexed on MeshKore from PyPI. Agent frameworks and tools published to the Python Package Index. Each entry links back to its project page.
25,155 agents · ranked by popularity · refine in the directory →
Source platform: PyPI →
PyPI agents — page 241 of 252
Legal AI agents framework — extension of smolagents with structured, jurisdiction-agnostic legal reasoning (French law example included)
Selective memory layer for AI agents: importance-scored, self-decaying, pluggable.
Ollama model components as a standalone Langflow Extension Bundle.
Haystack integration for Linkup Search API
Unofficial one-command catalog LLM deployment CLI for Linode
LiteLLM / OpenAI-compatible gateway doctor CLI (ping, tokens, context bench)
Custom LiteLLM providers for Pollinations.ai and AI Horde, two free image-generation APIs
TTTPS Proof-of-Time success callback for LiteLLM: cryptographic audit-trail timestamps for LLM completions via the public KPP Provenance API
Pareto-optimal OpenRouter provider chooser for LiteLLM
Provider-agnostic prompt optimization pipeline powered by LiteLLM.
Local proof harness for LiteLLM context-window routing
Sage - a terminal AI assistant for OpenAI-compatible API endpoints
Universal LLM API client for 165 providers — chat, streaming, tools, embeddings, search, OCR, plus an OpenAI-compatible proxy and an MCP server.
LangGraph BaseStore adapter for Lithtrix agent memory
Alias stub for agents-live (take your agents live). Install agents-live instead.
Haystack 2.x integration for the Live Tennis API: live scores, matches and players as Documents
Auto-tune and benchmark llama.cpp / ik_llama.cpp inference on NVIDIA, AMD (ROCm), and Apple Silicon GPUs
LlamaIndex integration for Common Compute — Apple-Silicon-powered LLMs and embeddings for your LlamaIndex pipeline.
llama-index llms iflytek spark integration
InfoLang memory integration for LlamaIndex
Structure-aware LlamaIndex node parser that turns xberg native chunks and elements into nodes
Perseus Vault persistent, local, encrypted memory for LlamaIndex — agent tools and a retriever backed by the Perseus Vault MCP engine.
llama-index postprocessor distil integration
LlamaIndex data loaders for Feishu/Lark wiki, docs, and drive
LlamaIndex readers for OpenDMA - integrate ECM systems with LlamaIndex
Official LlamaIndex reader for the WellMarked API — load any URL as clean Markdown.
LlamaIndex reader for 101 document formats powered by xberg's Rust extraction engine
LlamaIndex retrievers for OpenDMA - retrieve ECM content with LlamaIndex
LlamaIndex tool spec for chDB, the in-process ClickHouse engine: analytical SQL over local files, object storage, and remote databases with engine-level read-only safety.
llama-index tools ilovevideoeditor integration
Nimble Web Search tool for LlamaIndex
LlamaIndex tool for SERPdive, the AI Search API and Tavily alternative: answer-ready web content for agents, same speed, 20.2% fewer tokens, higher answer quality (60.7% of decided duels) on a public benchmark
Xpoz social media intelligence tools for LlamaIndex agents — Twitter/X, Instagram, Reddit, and TikTok
LlamaIndex VectorStore for MySQL 9's native VECTOR type — works with ShannonBase, self-hosted MySQL, and MySQL HeatWave.
Benchmark llama.cpp models across candidate configurations
Official Python SDK for LlamaGen Comic API and Animation API
TTTPS Proof-of-Time callback handler for LlamaIndex: cryptographic audit-trail timestamps for LLM events via the public KPP Provenance API
LlamaIndex web reader that loads URLs through ProxyHat residential proxies — rotation, geo-targeting, sticky sessions.
LlamaIndex tools and a mandatory pre-execution gate for RelayShield's MCP registry risk and prompt-injection breach checks.
LlamaIndex tools for U.CASH agent monetization over HTTP-402.
Benchmark the LLM inference capacity of a server (llama-benchy orchestrator with auto-detection, real-workload sizing, charts and reports).
Lightweight Python framework for building LLM agents with tool calling and RAG
A dev tool for manually driving llm-agents-from-scratch's SupervisedTaskHandler one call at a time, over HTTP, via a React frontend.
A resilient multi-provider LLM layer that automatically fails over to alternative models and providers when requests fail.
LaC (LLM as Code) - reference engine. Declarative behavior for LLM apps: law in files, perimeter in code.
Tamper-evident audit trails for LLM lifecycles: training, deployment, and monitoring.
Cross-vendor LLM quota / budget self-check CLI. Stable 0/1/2/3 exit-code contract lets any AI agent gate its work before burning through budget. Supports MiniMax Token Plan, OpenRouter credits, self-hosted LiteLLM proxy spend, and self-logged JSONL (any LLM vendor).
Cage: run local-LLM components in a network-isolated OS process, hardware access intact, with a friction-free call interface back to the rest of your app.
Free-first LLM routing for Python: local Ollama -> OpenRouter free models -> paid Anthropic API, with automatic fallback and a savings ledger.
Model names, pricing, and free-tier metadata for OpenAI, Anthropic, and Google Gemini.
LLM plugin to serve an OpenAI Chat Completions API endpoint
A tiny, readable circuit breaker for LLM calls: hard budget caps that fail over to a free local model instead of raising or overspending.
OpenRouter LLM benchmark for repair-agent and validator-agent model selection.
Multi-format LLM API compliance testing tool — validates endpoints against OpenAI Chat, Open Responses, Anthropic, and Google GenAI specs.
Fit chat history into any model's context window: token estimation, sliding-window packing with a running summary, and overflow detection. Zero dependencies, no AI inside.
Composable pre-call and post-call hooks for LLM API calls: pricing, budgets, cost caps, rate limits, event log, observability.
Pack the most important text into a fixed LLM context-window token budget.
Dequantize compressed-tensors NVFP4 LLM checkpoints to dense safetensors (streaming, constant memory)
A lightweight Prometheus exporter for LLM eval metrics: faithfulness, semantic drift, cost, and CI/CD regression gating.
Accurate FLOPs, memory, latency, and energy estimation for Large Language Models.
Provider-agnostic middleware for LLM safety
The htop for LLM inference. Measured. Not guessed.
A multi-provider LLM-as-a-Judge consensus panel with chained criteria for concept matching validation.
Self-hosted LLM observability SDK for Python, LangChain, and LangGraph
Access tools from MCP servers as LLM tools
A personal, local-first memory store for LLM applications and agents
LLM plugin for the Meta AI API
Run-scoped LLM observability control plane backed by LiteLLM
Perplexity's Agent API as a model for Simon Willison's llm CLI.
Runtime prompt-injection defense middleware for LLM applications
Run multi-stage LLM workflows: DAGs of prompts and scripts with automatic output chaining
Local, cross-provider preflight checks for an LLM model switch
A GitHub Action and CLI that detects risky changes to LLM prompts and AI configuration before they ship to production
Resilient cascading LLM inference across multiple providers with failover, circuit breaking, and retry cooldowns
Client-side rate limiting for the Claude API. Never sees your API key.
Transparent local proxy that redacts private information from LLM requests and seamlessly restores it in responses
Cost-aware, latency-optimized routing for multiple LLM API providers
Flight recorder for LLM agents: trace every call, replay runs deterministically, diff runs to find where behavior diverged.
Zero-dependency scrubbers for LLM chat output — strip leaked JSON state-op fences, backend scaffold blocks, and tool-preamble meta-narration from player-facing text
Defensive pipeline for LLM apps: prompt injection, tool-call gating, output/PII guards
High-performance enterprise PII redaction and context preservation proxy for Large Language Models
Deterministic Go and Python sketches for high-cardinality LLM data.
Local smoke tests for live LLM models, latency, and cost
Track LLM token usage, enforce policies & budget limits, and trigger alerts — across OpenAI, Groq, OpenRouter, AWS Bedrock, and custom providers.
Per-query token ledger for LLM pipelines: record what the API reports, answer with one GROUP BY. Zero dependencies, dashboard included.
Local LLM token efficiency middleware — zero cloud, zero tracking, coaches developers on token waste in the terminal
Track LLM API costs, tokens, and latency to MySQL
LLM plugin for models hosted by UnoRouter
Lightweight terminal tool to check remaining credits / quota across LLM providers. Zero dependencies.
A general-purpose LLM verification framework: fine-grained reward + Pivot Preference Tournament best-of-N selection.
Async LLM router across AWS Bedrock, Google Gemini, and Groq providers
Deterministic workflow topology enforcement for LLM-powered systems.
LLM plugin for Z.AI / GLM models with Coding Plan support
LLM plugin for Z.AI / GLM models with Coding Plan support
LLM plugin for ZenMux.ai - OpenAI Chat Completions + Anthropic Messages API
Interactive LLM-driven automated algorithm design with evolutionary optimization
LLM-driven interactive hypothesis generation for SMEFT global fits
A lightweight, personal Claude Code-style agentic CLI for LOCAL LLMs (LM Studio).
Compose a language model from scratch — and learn how it works. PySide6 educational GUI + CLI.
Scored tooling katas, habit rules, and pre-baked warm-start sessions for LLM coding agents