Language

Python agents

56,634 Python AI agents indexed on MeshKore — the most complete public catalog, ranked by popularity and updated daily.

56,634 agents · ranked by popularity · refine in the directory →

Python agents — page 513 of 567

visualLLM

a language model with visual context.

visualchatgpt

💬 Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models

visualisation-with-llm

A library for data visualization with LLMs

visuallm

Visualization tool for various generation tasks on Language Models.

vital-agent-container-client

Vital Agent Container Client

vital-agent-container-sdk

Vital Agent Container SDK

vital-agent-eval-env

Vital Agent Eval Env

vital-agent-kg-utils

Vital Agent KG Utils

vital-agentbox

Secure sandboxed code execution and agent toolbox

vital-llm-cluster-mgr

Vital LLM Cluster Manager

vital-llm-reasoner

Vital LLM Reasoner

vitax-rag

ViTax-RAG: Alignment-augmented language model for viral taxonomy classification

vitrage

The OpenStack RCA Service

vitrage-dashboard

Vitrage Horizon plugin

vitrage-tempest-plugin

Tempest plugin for Vitrage project

vitrage-tempest-tests

Tempest plugin for Vitrage project

vkyGPT

A sample AI

vlagents

Modular VLA and Environment interfaces.

vlite-storage

Super simple vector data storage based on vectorlite

vllama

Comprehensive CLI tool and VS Code extension for vision models, AutoML, and local LLMs

vlllm

A utility package for text generation using vLLM with multiprocessing support.

vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-acc

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-agentic-apis

This name has been reserved using Reserver

vllm-agentic-stack

This name has been reserved using Reserver

vllm-allocator-adaptor

vLLM Allocator Adaptor (C/CUDA/Python) using callback shims

vllm-ascend

vLLM Ascend backend plugin

vllm-benchmark-suite

The most comprehensive benchmarking suite for vLLM inference servers

vllm-block

Block implementation of vLLM

vllm-bootstrap-client

gRPC client for vllm-bootstrap inference service

vllm-canon

vLLM plugin: out-of-tree registration of canon-layer architectures (e.g. LlamaCanonForCausalLM from PhysicsLM4)

vllm-cli

A CLI tool to conveniently serve LLMs with vLLM

vllm-client

Client for the vLLM API with minimal dependencies

vllm-cluster-manager

Deploy, manage, and monitor vLLM instances across a GPU cluster from a single web dashboard.

vllm-consul

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-cpm

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-cpu

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-cpu-amxbf16

vLLM CPU inference engine (AVX512 + VNNI + BF16 + AMX optimized)

vllm-cpu-avx512

vLLM CPU inference engine (AVX512 optimized)

vllm-cpu-avx512bf16

vLLM CPU inference engine (AVX512 + VNNI + BF16 optimized)

vllm-cpu-avx512vnni

vLLM CPU inference engine (AVX512 + VNNI optimized)

vllm-df11

Dfloat11 plugin for vLLM

vllm-doctor

Diagnostic tool for vLLM inference servers

vllm-dolphin
vllm-efficient-client

A unified interface for efficient LLM inference with vLLM and OpenAI-compatible APIs

vllm-emissary

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-factory

The LEGO set for custom vLLM model plugins — build, test, and deploy custom encoders, poolers, and kernels

vllm-fixed

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-flash-attn

Forward-only flash-attn

vllm-gguf-plugin

Out-of-tree GGUF quantization plugin for vLLM

vllm-haystack

Haystack integration for vllm

vllm-hpu-extension

HPU extension package for vLLM

vllm-htop

htop-style terminal monitor for vLLM inference servers

vllm-hust

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-inference

Inference locally.

vllm-iter

Iterable-based offline generation helpers for vLLM.

vllm-judge

LLM-as-a-Judge evaluations for vLLM hosted models

vllm-kunlun

vLLM Kunlun3 backend plugin

vllm-lens

vLLM plugin for interacting with activations during inference

vllm-logits

A collection of useful util functions

vllm-manager

Multi-instance vLLM cluster orchestration and log management

vllm-marconi-offload

Two-tier (RAM + SSD) KV cache offload connector for vLLM with Marconi-style reuse-aware eviction.

vllm-marenostrum

Super simple vLLM server launcher for SLURM/HPC with nested config support

vllm-mbart
vllm-mcp-server

MCP server for vLLM - expose vLLM capabilities to AI assistants

vllm-messages

This name has been reserved using Reserver

vllm-metal

vLLM hardware plugin for Apple Silicon - unifies MLX and PyTorch under a single lowering path

vllm-mindspore

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-mini

vLLM mini.

vllm-mlx

vLLM-like inference for Apple Silicon - GPU-accelerated Text, Image, Video & Audio on Mac

vllm-mock

Provide mock instance to test vllm without CUDA or any GPUs.

vllm-mon

Production-grade vLLM metrics monitoring TUI with persistent storage and Grafana-style visualizations

vllm-musa

vLLM platform plugin for Moore Threads MUSA GPUs

vllm-nccl-cu11
vllm-nccl-cu12
vllm-npu

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-omni

A framework for efficient model inference with omni-modality models

vllm-online

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-playground

A web interface for managing and interacting with vLLM servers

vllm-plugin-meralion2

A vLLM plugin to register the MERaLiON-2-10B model architecture with vLLM’s plugin system.

vllm-rbln

vLLM plugin for RBLN NPU

vllm-recorder
vllm-responses

This name has been reserved using Reserver

vllm-rocm

A high-throughput and memory-efficient inference and serving engine for LLMs with AMD GPU support

vllm-router

High-performance Rust-based load balancer for VLLM with multiple routing algorithms and prefill-decode disaggregation support

vllm-rs

A minimal, high-performance large language model (LLM) inference engine implementing vLLM in Rust.

vllm-sdk

Minimal Python SDK for the vLLM API

vllm-semantic-router-bench

Comprehensive benchmark suite for semantic router vs direct vLLM evaluation across multiple reasoning datasets

vllm-speculative-autoconfig

Automatic configuration planner for vLLM - Eliminate the guesswork of configuring vLLM by automatically determining optimal parameters

vllm-spyre

vLLM plugin for Spyre hardware support

vllm-spyre-next

Next iteration of vllm-spyre on the torch-spyre stack

vllm-sr

vLLM Semantic Router - Intelligent routing for Mixture-of-Models

vllm-sr-sim

vLLM Semantic Router fleet simulator for capacity planning, SLO validation, and what-if analysis

vllm-swift

vLLM Metal plugin powered by mlx-swift — high-performance LLM inference on Apple Silicon

vllm-test-tpu

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-tgis-adapter

vLLM adapter for a TGIS-compatible grpc server

vllm-top

A monitoring tool for vLLM metrics.

vllm-tpu

A high-throughput and memory-efficient inference and serving engine for LLMs

vllm-tuner

A Python package for tuning vLLM hyperparameters.

vllm-usf

vLLM-USF: A high-throughput and memory-efficient inference engine for LLMs (USF Custom Build)

Browse other language pages