Category
Image agents
2,727 Image AI agents indexed on MeshKore — the most complete public catalog, ranked by popularity and updated daily.
2,727 agents · ranked by popularity · refine in the directory →
Image agents — page 27 of 28
Face Shape Detector AI review: compare features, pricing, use cases, access model, and alternatives for this AI Detection agent in 2026. Find your celebrity lookalike using AI with just a photo.
Generate an Open Knowledge Format (OKF) bundle from your docmd site for AI agent consumption.
n8n community node for SiliconFlow (硅基流动). Zero runtime dependencies. Provides a SiliconFlow action node (Chat / Vision / Embeddings / Image / Rerank / Audio TTS+ASR / Video) and a LangChain-compatible Chat Model node for AI Agents. Installs cleanly witho
Put an OpenAI Realtime voice agent on Microsoft Teams calls. Terminates the StandIn media bridge wire protocol on one side and the OpenAI Realtime API WebSocket on the other: barge-in, function tools, on-demand vision, call governors. PCM 16k on the wire,
Official JavaScript/TypeScript SDK for inference.sh - Run AI models with a simple API
Agentic MCP server + quality workbench that records, generates, and maintains Playwright UI tests for any web project.
Turn any machine or Google Colab notebook into a one-line AI model server with an auto-generated public API.
AG-UI integration for the Anthropic Claude Agent SDK, with multimodal attachment support (images and documents)
A way of controlling deployed agents to keep your (and their) sanity.
Python and gRPC stubs generated by Alis Build for alis.open.agent.v1
Python and gRPC stubs generated by Alis Build for alis.open.agent.v2
Generate agent tools (and an MCP server) from your API — then verify each tool against the live API, repair what breaks, and grade whether a real agent can finish real tasks.
Tool to generate a changelog from changelog fragments
A proxy service converts ModelScope and SiliconFlow image APIs to OpenAI-compatible format
Multi-agent CLI that generates Terraform and deploys full-stack apps to AWS/GCP free tier
Custom LiteLLM providers for Pollinations.ai and AI Horde, two free image-generation APIs
Official Python SDK for LlamaGen Comic API and Animation API
Generate and validate llms.txt files — CLI and Python library, zero dependencies.
Multimodal knowledge base for documents and images with intent-aware retrieval, grounded answers, and evidence images.
Audit any OpenAPI spec for AI-agent and MCP readiness, and generate an MCP server scaffold.
CrewAI tools for Provision Stack — deploy cloud infrastructure via AI agents.
LangChain tools for Provision Stack — deploy cloud infrastructure via AI agents.
A practical GUI and Python toolkit for Pixeltable YOLOX
Raven — AI-assisted learning companion. Auto-detect materials, split by chapters, generate quizzes, organize notes.
AI-powered multi-agent framework to plan, generate and heal Selenium Python tests (pytest + pytest-bdd)
Hybrid vision web agent — planner + UI-TARS grounder + Playwright
Companion plugin: load & serve pre-quantized humming Qwen-Image DiT pipelines on stock vLLM-Omni
Photo to Text Converter review: compare features, pricing, use cases, access model, and alternatives for this Productivity agent in 2026. Extract editable text from images and PDFs online
Was it AI review: compare features, pricing, use cases, access model, and alternatives for this AI Detection agent in 2026. Free tool to detect if an image was generated by artificial intelligence.
AI Image Combiner review: compare features, pricing, use cases, access model, and alternatives for this Images agent in 2026. AI tool for merging two images into new creative scenes
OnlyGFs.ai review: compare features, pricing, use cases, access model, and alternatives for this NSFW agent in 2026. AI girlfriend chatbot with memory, image generation, and roleplay.
RenderPop review: compare features, pricing, use cases, access model, and alternatives for this Images agent in 2026. Free AI image and video generator requiring no sign-up.
Wan 3.0 AI review: compare features, pricing, use cases, access model, and alternatives for this AI Video Agents agent in 2026. Free online AI tool for text-to-video and image-to-video generation.
Lynote helps you rewrite AI-generated text to sound like a real person wrote it. It fixes awkward phrasing and repetitive patterns while keeping your original m
Consent-based local observer for Phigon AI agent supervision.
Browser multiplayer SDK for games generated by Loova AI Agent
Rulvar isolated tool executors: reference ToolExecutorProvider adapters that run tool work out of process (subprocess and container) so hostile or model-generated scripts cannot reach host capabilities.
Render CLI LaTeX and read Markdown with images directly in the terminal.
Shared invoke-boundary core for agent-governance-demo: host adapters, entitlement/policy resolution, the governance envelope, and the registry Postgres schema. Published to PyPI via OIDC Trusted Publishing (.github/workflows/publish.yml) and installed into registry-api/workflows images.
Project lifecycle CLI for agentive projects — task status flow, environment doctor, evaluator provisioning
Hierarchical AI-generated index over a claudesync export of your claude.ai conversations.
CLI that prepares local images for manual agent input, plus an offline image variant benchmark harness.
LangChain document loader that turns PDFs and images into Markdown documents with the HexRead API.
LlamaIndex reader that turns PDFs and images into Markdown documents with the HexRead API.
Type annotations for boto3 AgentRegistry 1.43.85 service generated with mypy-boto3-builder 8.12.0
Type annotations for boto3 AgentRegistryControl 1.43.84 service generated with mypy-boto3-builder 8.12.0
LLM-assisted OpenModelica modeling: a tested agentic generate-compile-simulate-verify loop with structured omc diagnostics, quantitative trajectory verification, and a benchmark task ladder.
Type annotations for boto3 AgentRegistry 1.43.85 service generated with mypy-boto3-builder 8.12.0
Type annotations for boto3 AgentRegistryControl 1.43.66 service generated with mypy-boto3-builder 8.12.0
OhAPI review: compare features, pricing, use cases, access model, and alternatives for this NSFW agent in 2026. The world’s leading full-stack NSFW AI API: chat, voice, image, video, live cams & Digital Twins.
TextileGen review: compare features, pricing, use cases, access model, and alternatives for this Images agent in 2026. AI-powered tool for designing textile patterns and fabric prints.
HiAPI review: compare features, pricing, use cases, access model, and alternatives for this AI Agents Platform agent in 2026. One API for leading AI image, video, and text generation models.
Overchat is a simple platform that brings together over 50 of the world's best AI models. You can chat, write, edit images, and create videos all in one secure
MCP server for pulr.ai — secure FILE exchange infrastructure for AI agents: upload, share, fetch, and revoke durable, verified, expiring file artifacts (reports, CSVs, images, PDFs) with provenance, scanning, and audit. Backend file handoff — not a UI str
AEP 0.1 SDK: emit/consume/control helpers over the schema-generated envelope and payload types
Companion intelligence and real-time decision supervision for autonomous AI agents.
Speechify TTS integration for Vision Agents
ArchOne AI review: compare features, pricing, use cases, access model, and alternatives for this Web AI Agents agent in 2026. All-in-one AI spatial design visualization platform for interior, exterior and landscape design.
LongTerMemory review: compare features, pricing, use cases, access model, and alternatives for this Education agent in 2026. Study everything with custom and scheduled study plan and auto-generated Q&A pairs
Image to 3D AI creates 3D models from uploaded JPG, PNG, or WebP images, or from text prompts. Choose textured or geometry-only output, preview it online, and e
TypeScript types + client for the ai-agent-subsystem CRDs (Agent / Station / AgentDefinition). Types are generated from the D source so they cannot drift.
Multi-provider language generation toolkit using Vercel AI SDK - generate responses with 10+ LLM providers including OpenAI, Anthropic, Google, and more.
LLM-agnostic AI platform for SAP CAP. 11 providers behind one API (Anthropic, Google Gemini, AWS Bedrock, Azure OpenAI, Ollama, Groq, Fireworks, DeepSeek, Mistral, OpenAI-compatible, SAP Generative AI Hub) with 30+ middleware primitives spanning cost/resi
Improve OpenAI in pi with fast mode, usage stats, realtime voice, image generation, and footer polish.
OpenAI SDK adapter for llm-ports. Implements LLMPort and EmbeddingsPort. baseURL support covers 10+ OpenAI-compatible providers (Azure, Groq, Together, Fireworks, DeepInfra, Perplexity, Cerebras, LiteLLM proxy).
Deterministic policy engine for LLM-generated SQL: SELECT-only, PII column denylist, table allowlist, cost cap. Multi-dialect via sqlglot (BigQuery, Snowflake, Postgres, Trino, …).
Stream-level supervision for AI agents: watch live token streams, interrupt mid-generation, gate tool actions pre-execution. Name reservation - the open-source release lands here in September 2026.
Run LangGraph graphs as errand jobs — background execution, HITL resume, SSE streaming, smart retries, and an auto-generated FastAPI router.
MCP server for image generation, editing and multi-turn refinement with OpenAI gpt-image-2.5
Host-level deployment and supervision agent for the MCP Worker platform: manage local Worker services (git/pip/systemd/launchd), health polling, heartbeat aggregation, and self-healing.
Meshy helps you make 3D models by chatting, using text, or starting from images. It can prepare models for printing, add textures, rig characters, and export fi
DeepSeek Harness 外置视觉模型插件:inspect_image 把本地图片或 http(s) 图片 URL 发给任意 OpenAI 兼容端点,视觉模型看图的文字回答直接带回对话;附 Web UI 设置栏。
多模态视觉 MCP + 文生图 Skill 工具包 (OpenAI GPT-4o · 通义千问 Qwen-VL · Google Gemini · Anthropic Claude)
Spec-driven SDK for the gptproto async media generation API
LangChain tools for Magic Hour: AI text-to-video, image-to-video and image generation (Sora 2, Veo 3.1, Kling 3.0, WAN 2.2, GPT-image, Nano Banana Pro, ...)
llama-index tools magic_hour integration (AI video and image generation)
Image Background Changer review: compare features, pricing, use cases, access model, and alternatives for this Images agent in 2026. Free tool to instantly remove Bg From Images
Video2Jpg review: compare features, pricing, use cases, access model, and alternatives for this Images agent in 2026. Free Online Video to Image Converter (JPG, PNG, WebP)
Bg Remover review: compare features, pricing, use cases, access model, and alternatives for this Images agent in 2026. Instant, precise background removal for images and photos.
A generic MCP server that lets a text-only LLM agent 'see' images by sending them to any OpenAI-compatible multimodal model (e.g. gpt-4o, glm-4.6v, qwen-vl).
Image to Layers AI review: compare features, pricing, use cases, access model, and alternatives for this Images agent in 2026. AI that splits any flat image into editable transparent layers (PNG/PSD)
H3 Max Turbo review: compare features, pricing, use cases, access model, and alternatives for this AI Video Agents agent in 2026. Fast, browser-based text-to-video and image-to-video generation.
Stivio turns a photo into a 3–30 second HD video. Upload an image, describe the motion in plain English, choose a model, and download the finished MP4 in minute
让任意 AI Agent 直接调用 focalapi 创作模型的命令行工具
An Antora extension to generate llm.txt files for LLM consumption.
Magic Hour tools for LangChain.js — AI video (Sora 2, Veo 3.1, Kling 3.0, Seedance, WAN 2.2) and image generation as LangChain tools.
Continuous realtime voice for Hermes Agent. Talk while Hermes runs background tasks.
Generated UTDK provider client for OpenAI API. The OpenAI REST API. Please see https://platform.openai.com/docs/api-reference for more details.
Multi-agent orchestrator SoT — deep-merge rulesync trees, product scripts, and generate pipeline into a consumer repo
Image workflows for the Glove agent framework — prompt pipelines with enhancer inbetweens, first-class characters and scenes, reference images and assembly, behind a BYO image-model adapter.
Unified observability for LLM/text, image, and video calls (Django adapter)
ARK Runtime SDK — attach ARK decision telemetry, cost attribution, and (experimental) constrained supervision around your existing agent.
Read-only CLI and TUI over the generated llms-explorer concept tree and llms-concept-abstractor concept packs, plus a thin CLI over the skills SDK
CreateForge AI review: compare features, pricing, use cases, access model, and alternatives for this AI Video Agents agent in 2026. AI image and video creation workspace for creators and teams
Autonomous API Schema & Reliability Inspector CLI. Probes OpenAPI specs, generates test suites, executes real-time HTTP calls, and exports Postman collections.
MCP server for editable GIMP 3 workflows: live previews, grouped edits, PDB and GEGL discovery, pixel measurement and recipes
One client for every OpenAI-compatible LLM server - local (Ollama, LM Studio, vLLM, llama.cpp) or hosted (OpenRouter, OpenAI) - with images and audio as plain message parts.
GPT Image 2.5 review: compare features, pricing, use cases, access model, and alternatives for this Images agent in 2026. Advanced AI image generation and editing model with different API tiers.
Seedance 2.0 AI Video Maker review: compare features, pricing, use cases, access model, and alternatives for this Ecommerce agent in 2026. Multimodal AI video generation with text, image, video, and audio references.
Open video production system for humans and agents — local-first CLI for directable, reproducible video: scene scripts, deterministic rendering, local speech, word-synced captions, product walkthroughs, AI clips, revisions, and provenance.