Language
Python agents
56,634 Python AI agents indexed on MeshKore — the most complete public catalog, ranked by popularity and updated daily.
56,634 agents · ranked by popularity · refine in the directory →
Python agents — page 513 of 567
a language model with visual context.
💬 Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models
A library for data visualization with LLMs
Visualization tool for various generation tasks on Language Models.
Vital Agent Container Client
Vital Agent Container SDK
Vital Agent Eval Env
Vital Agent KG Utils
Secure sandboxed code execution and agent toolbox
Vital LLM Cluster Manager
Vital LLM Reasoner
ViTax-RAG: Alignment-augmented language model for viral taxonomy classification
The OpenStack RCA Service
Vitrage Horizon plugin
Tempest plugin for Vitrage project
Tempest plugin for Vitrage project
A sample AI
Modular VLA and Environment interfaces.
Super simple vector data storage based on vectorlite
Comprehensive CLI tool and VS Code extension for vision models, AutoML, and local LLMs
A utility package for text generation using vLLM with multiprocessing support.
A high-throughput and memory-efficient inference and serving engine for LLMs
A high-throughput and memory-efficient inference and serving engine for LLMs
This name has been reserved using Reserver
This name has been reserved using Reserver
vLLM Allocator Adaptor (C/CUDA/Python) using callback shims
vLLM Ascend backend plugin
The most comprehensive benchmarking suite for vLLM inference servers
Block implementation of vLLM
gRPC client for vllm-bootstrap inference service
vLLM plugin: out-of-tree registration of canon-layer architectures (e.g. LlamaCanonForCausalLM from PhysicsLM4)
A CLI tool to conveniently serve LLMs with vLLM
Client for the vLLM API with minimal dependencies
Deploy, manage, and monitor vLLM instances across a GPU cluster from a single web dashboard.
A high-throughput and memory-efficient inference and serving engine for LLMs
A high-throughput and memory-efficient inference and serving engine for LLMs
A high-throughput and memory-efficient inference and serving engine for LLMs
vLLM CPU inference engine (AVX512 + VNNI + BF16 + AMX optimized)
vLLM CPU inference engine (AVX512 optimized)
vLLM CPU inference engine (AVX512 + VNNI + BF16 optimized)
vLLM CPU inference engine (AVX512 + VNNI optimized)
Dfloat11 plugin for vLLM
Diagnostic tool for vLLM inference servers
A unified interface for efficient LLM inference with vLLM and OpenAI-compatible APIs
A high-throughput and memory-efficient inference and serving engine for LLMs
The LEGO set for custom vLLM model plugins — build, test, and deploy custom encoders, poolers, and kernels
A high-throughput and memory-efficient inference and serving engine for LLMs
Forward-only flash-attn
Out-of-tree GGUF quantization plugin for vLLM
Haystack integration for vllm
HPU extension package for vLLM
htop-style terminal monitor for vLLM inference servers
A high-throughput and memory-efficient inference and serving engine for LLMs
Inference locally.
Iterable-based offline generation helpers for vLLM.
LLM-as-a-Judge evaluations for vLLM hosted models
vLLM Kunlun3 backend plugin
vLLM plugin for interacting with activations during inference
A collection of useful util functions
Multi-instance vLLM cluster orchestration and log management
Two-tier (RAM + SSD) KV cache offload connector for vLLM with Marconi-style reuse-aware eviction.
Super simple vLLM server launcher for SLURM/HPC with nested config support
MCP server for vLLM - expose vLLM capabilities to AI assistants
This name has been reserved using Reserver
vLLM hardware plugin for Apple Silicon - unifies MLX and PyTorch under a single lowering path
A high-throughput and memory-efficient inference and serving engine for LLMs
vLLM mini.
vLLM-like inference for Apple Silicon - GPU-accelerated Text, Image, Video & Audio on Mac
Provide mock instance to test vllm without CUDA or any GPUs.
Production-grade vLLM metrics monitoring TUI with persistent storage and Grafana-style visualizations
vLLM platform plugin for Moore Threads MUSA GPUs
A high-throughput and memory-efficient inference and serving engine for LLMs
A framework for efficient model inference with omni-modality models
A high-throughput and memory-efficient inference and serving engine for LLMs
A web interface for managing and interacting with vLLM servers
A vLLM plugin to register the MERaLiON-2-10B model architecture with vLLM’s plugin system.
vLLM plugin for RBLN NPU
This name has been reserved using Reserver
A high-throughput and memory-efficient inference and serving engine for LLMs with AMD GPU support
High-performance Rust-based load balancer for VLLM with multiple routing algorithms and prefill-decode disaggregation support
A minimal, high-performance large language model (LLM) inference engine implementing vLLM in Rust.
Minimal Python SDK for the vLLM API
Comprehensive benchmark suite for semantic router vs direct vLLM evaluation across multiple reasoning datasets
Automatic configuration planner for vLLM - Eliminate the guesswork of configuring vLLM by automatically determining optimal parameters
vLLM plugin for Spyre hardware support
Next iteration of vllm-spyre on the torch-spyre stack
vLLM Semantic Router - Intelligent routing for Mixture-of-Models
vLLM Semantic Router fleet simulator for capacity planning, SLO validation, and what-if analysis
vLLM Metal plugin powered by mlx-swift — high-performance LLM inference on Apple Silicon
A high-throughput and memory-efficient inference and serving engine for LLMs
vLLM adapter for a TGIS-compatible grpc server
A monitoring tool for vLLM metrics.
A high-throughput and memory-efficient inference and serving engine for LLMs
A Python package for tuning vLLM hyperparameters.
vLLM-USF: A high-throughput and memory-efficient inference engine for LLMs (USF Custom Build)