Category
Image agents
2,343 Image AI agents indexed on MeshKore — the most complete public catalog, ranked by popularity and updated daily.
2,343 agents · ranked by popularity · refine in the directory →
Image agents — page 4 of 24
DALLE2 in the command line.
InterfaceAgent: a versatile framework designed to create system and interface agents capable of managing mobile and desktop applications and features.
Muse Studio is an Agentic AI workspace for planning, visualizing, and iterating on stories and video creation.
TaskFlowAI is a lightweight and flexible framework designed for creating AI-driven task pipelines and multi-agent workflows. It provides developers with a streamlined approach to building agentic systems
智元 IIM 是一款开源的网页版即时聊天系统, 同时拥有AI聊天对话功能, 支持ChatGPT、Midjourney、文心一言、讯飞星火、通义千问等AI助手功能
Implementation of "MADiff: Offline Multi-agent Learning with Diffusion Models"
AI image generation and editing via Gemini, OpenAI, Fal and Replicate, right in your WordPress media library. Native-like integration in Elementor, WooCommerce, and other plugins.
AI-powered assistant for EasyEDA — generate schematics from natural language, browse LCSC components, design PCBs with custom DRC configurations, and get interactive circuit design help.
TerraMours实战项目,基于vue3.0+ts+naive UI+vite的ChatGPT项目前端。实现用户登陆和基于SK的多语言模型聊天、基于chatgpt和SD的多模型图片生成等功能。
ArchGuard Co-mate is an AI-powered architecture copilot, design and governance tools.
Private local AI Photographer Agent on AMD Radeon and ROCm
Zero Graph – Minimalist LLM framework designed for AI Agent programming
InsightSolver: Colab notebooks for exploring and solving operational issues using deep learning, machine learning, and related models.
A powerful Python framework for orchestrating AI agents and managing complex LLM-driven tasks with ease.
Multi-Modal-AI-Orchestrator (Reset version),AI Full-modal Full-agent:Text → Image → Music → Lights → Video, Includes "Scenario Director, Parent-Child Theater, Biofeedback, Party Collaboration Station, Scenario Recipes" - five ready-to-use scenario packages. AI全模态+全agent:含 “情境导演、亲子剧场、生物反馈、派对协作台、情境食谱” 五大开箱即用场景
Speech o Text using docker image with ggerganov/whisper.cpp
Libriscribe is an AI-powered open-source book creation system. The system is designed to help you write your book, from initial concept to a polished manuscript. It uses a multi-agent approach, where specialized AI agents collaborate to handle different stages of the book creation process.
Promptdesk is a tool designed for effectively creating, organizing, and evaluating prompts and large language models (LLMs).
Self-hostable, open-source alternative to Claude Managed Agents. Multi-LLM support (Anthropic, OpenAI, Ollama), enterprise governance (org/teams/RBAC)
A docker image for running AI agents in YOLO mode
A full featured, powerful, and efficient AI Agentic Harness/Desktop Agent app designed from the ground up with local inference on consumer hardware in mind. It evolves, grows, and gets smarter as you go. And it REMEMBERS...
Stitch 提出的设计系统文档格式 —— 用纯文本 Markdown 记录设计系统,让 AI 编码 Agent 能够生成风格一致的 UI。
An advanced LangGraph series exploring real-world agent workflows, dynamic tools, parallel execution, long-term memory, and human-in-the-loop designs. Includes hands-on Python notebooks for building scalable, production-ready AI agent architectures.
gradio region:us
The AVR Infrastructure project is designed to launch the Agent Voice Response application, which will start the Core, ASR, LLM, and TTS services integrated with Asterisk Audiosocket.
Devr.AI is an advanced AI-powered Developer Relations (DevRel) assistant designed to revolutionize open-source community management. By integrating with platforms like Discord, Slack, GitHub, and Discourse, Devr.AI functions as a virtual DevRel advocate that helps maintainers engage with contributors and streamline onboarding processes.
Fast, self-hostable sandboxes for agents
Self-designing, durable, self-improving agent orchestration: describe a goal — an agent designs the graph, a fleet of sealed full-agent nodes runs it, and a learning loop makes it better every run.
This repository contains an application using ROS2 Humble, Gazebo, OpenAI Gym and Stable Baselines3 to train reinforcement learning agents for a path planning problem.
一个用于在Unity中开发在线和离线聊天机器人的管线。A pipeline designed to create online and offline chat-bot in Unity.
Teams AI Bot integrated with several LLMs services (ChatGPT, GPT-3, DALL-E) from Azure OpenAI & OpenAI. Support Teams Message Extensions, Teams Task Modules and Adaptive Cards.
Photoshop for agents
A modern AI chatbot with chat, image generation, and text-to-speech features, designed for a smooth and friendly user experience.
FlowSteer: agents designing agentic workflows via reinforced progressive canvas editing.
Design Patterns for Multi Agents Frameworks Like Autogen, Langraph, Taskweaver,Crewai,etc
Negotiation Multi-Agent System (A negotiation library designed for situated negotiations within business-like simulations)
Open-source 3D AI agent framework — GLB/glTF avatars with LLM brains, memory, emotions, and autonomous payments. MCP server · x402 · Solana/EVM · Three.js. Embed anywhere as a web component. Character studio, animation gallery, OAuth 2.1. Browser-native.
A Swift library for macOS automation — mouse, keyboard, screenshots, image recognition, and AI-powered agents.
The KCORES Agent benchmarking project is designed to evaluate the tool-call capabilities of single-modal/multi-modal models.
[ICML 2026 Oral] Photoagent: A fully automated, intelligent photo-editing agent that autonomously plans multi-step aesthetic enhancements, smartly chooses diverse editing tools, and enables everyday users to achieve professional-looking results without crafting complex prompts.
Local-first floating desktop AI assistant for Windows — selection actions, OCR, notes, and a tool-calling agent.
LLM Agentic Tool Mesh Platform is an innovative platform designed to streamline and enhance the use of AI in various applications. It serves as a central hub to orchestrate 'Intelligent Plugins,' optimizing AI interactions and processes.
Dartantic is an agentic framework designed to make building client and server-side apps in Dart with generative AI easier and more fun!
An open-source project designed to provide an educational and collaborative AI-powered platform
A curated collection of resources, frameworks, papers, and best practices for designing, evaluating, and deploying agentic AI systems—from architecture patterns and safety considerations to real-world applications.
An AI-app that allows you to upload a PDF and ask questions about it. It uses StableVicuna 13B and runs locally.
CanvasAnvil is an AI multi-canvas creation platform for flowcharts, interior design, presentations, posters, infographics, and product storytelling.
Agent wallet infrastructure — encrypted keys, policy enforcement, credential proxy, auth platform. Self-hostable, multi-tenant, open source.
MedSAM-Agent: Empowering Interactive Medical Image Segmentation with Multi-turn Agentic Reinforcement Learning
The Retail Shopping Assistant is an AI-powered blueprint that provides a comprehensive interface for an intelligent retail shopping advisor. Built with LangGraph for agent orchestration, it features multi-agent architecture, real-time streaming responses, image-based search, and intelligent shopping cart management.
The video editor built for AI. Create videos with React + TypeScript — designed for AI agents, LLM pipelines, and automated production. Fully open source.
A collection of Agent Skills by Cali Castle.
Create amazing Stable Diffusion prompts with minimal prompt knowledge. A vicuna based prompt engineering tool for stable diffusion
An open source chat bot architecture for voice/vision (and multimodal) assistants, local(CPU/GPU bound) and remote(I/O bound) to run.
GPT-3 client for Windows and Unix with memories management that supports both text and speech in any language. Includes a free text2image
A simple matrix bot that supports image generation and chatting using ChatGPT
Agent skill for simulating gmoverid characteristics and gmoverid-based design
🐧 Unify-Agent: An end-to-end unified multimodal agent for faithful, knowledge-grounded image generation.
transformers safetensors gemma4 image-text-to-text bf16 multimodal
This app uses the OpenAISwift library, ChatGPTSwift library and OpenAI library to communicate with the popular ChatGPT artificial intelligence. The app allows you to have a quick message exchange with a simple and clean interface, but with useful features.
On-device AI SDK for Flutter — LLM inference, vision, STT, TTS, image generation, embeddings, RAG, and function calling. Metal GPU on iOS/macOS.
565 AI-callable tools across 16 MCP servers. Full-pipeline AAA game asset production. Controls Blender, Substance Suite, Maya, Houdini, and Unreal Engine 5. 50 specialized AI agents. One prompt in, game-ready asset out.
Open-source, local-first AI App Store screenshot editor.
Secure storage designed for Hyperledger Aries agents.
A Straightforward, Step-by-Step Implementation of a Video Diffusion Model
🦞 Production-grade AI Agent design system — 龙虾教练:产品级 AI Agent 设计体系
Radxa SoM carrier board design files + agentic hardware-db
Design Multi-Agent AI Systems Using MCP and A2A, by Packt Publishing
The open-core AI workbench — notebooks, agents, RAG, voice, and images across any model: OpenAI, Anthropic, Google, xAI, or local via Ollama/vLLM. BSL 1.1, auto-converting to Apache-2.0 on a two-year clock. Your AI keeps running when theirs doesn't.
AI4U is a plugin that allows you use the Godot Game Engine to specify agents with reinforcement learning. Non-Player Characters (NPCs) of games can be designed using ready-made components.
Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation
AI生图工作台 (AI Image Generator)、视频工作台。支持 Agent 与工作台模式、无限画布、素材管理、提示词反推/广场及 PWA,支持外链式配置;可接入 Grok、GPT Image 2 等自定义模型生图。自适应手机端。
Docker image for OMS (Operations Management Suite) Linux agent.
gradio mcp-server region:us
A standardized, application-agnostic narrative structure schema designed for reliably transporting authorial intent across multi-agentic narrative systems.
Schola is a plugin for enabling Reinforcement Learning (RL) in Unreal Engine. It provides tools to help developers create environments, define agents, and connect to python-based RL frameworks such as OpenAI Gym, RLlib or Stable Baselines 3 for training agents with RL.
Design with Intent: A collection of specialized AI agents and skills for experience design and strategy.
28 hand-designed 5x5 dot-matrix loader animations as standalone animated SVGs and a React component. ~4KB each, no JS runtime, agent-themed patterns.
Design Patterns for Multi Agents Frameworks Like Autogen, Langraph, Taskweaver, Crewai,etc
The MVC framework for chat apps built on a platform designed for communication systems.
Awesome Multimodal Assistant is a curated list of multimodal chatbots/conversational assistants that utilize various modes of interaction, such as text, speech, images, and videos, to provide a seamless and versatile user experience.
AmigaOS 3.1/4.1 and MorphOS application for chatting with ChatGPT or generating images
ExploitGym is a large-scale, realistic benchmark built from real-world vulnerabilities designed to evaluate AI agents' ability to develop exploits.
Generate AI-powered videos and images from the terminal using the `agent-media` CLI.
🧬 KIP is a Knowledge-memory Interaction Protocol designed for LLMs, aiming to build a sustainable learning and self-evolving memory system for AI Agents.
🌼 A token-friendly local MCP server for DaisyUI component documentation using their public llms.txt.
An open-source CLI for deploying LangChain agents to iMessage in seconds.
A Multipurpose GPT Chatbot with AI Image Generator, Counting , Logging, Leveling, Reaction Roles, Boost Tracker, Welcomer, Ticket.
A beginner-friendly and extensible Agentic RAG project that demonstrates the full pipeline of document parsing, retrieval, reranking, workflow orchestration, tool calling, and answer generation, designed for both learning and secondary development.
面向 AI Agent 的 AIGC 漫剧视频创作全流程工具集。
A demo app showcasing Vector Search using Azure AI Search, Azure OpenAI for text embeddings, and Azure AI Vision for image embeddings.
[CVPR 2025] MAC-Ego3D: Multi-Agent Gaussian Consensus for Real-Time Collaborative Ego-Motion and Photorealistic 3D Reconstruction
Agent-agnostic PPT design skills for creating high-end, editable presentation styles.
Opinionated UI constraints extracted from the best design systems. Use with AI agents to build pixel-accurate interfaces.
streamlit region:us
A task management system designed for AI development
A multi-agent reinforcement learning environment to design and benchmark control strategies aimed at reducing drag in turbulent open channel flow
LLM-driven browser automation library built on Playwright with 67 CLI/SDK tools, stable snapshot refs, and stealth mode.基于 Playwright 的 LLM 驱动浏览器自动化库,提供 67 个 CLI/SDK 工具、稳定快照引用(ref)与默认隐身模式,适用于 AI Agent 端到端网页操作。
🏛️ Directive · OpenClaw Multi-Agent Orchestration System — 10 AI agents modeled after the U.S. Federal Executive Branch. Dual independent veto layers (White House Counsel + OMB), per-department Inspector General auditing, inter-agency task forces, and a real-time Situation Room dashboard. Separation of powers, by design.
transformers safetensors qwen3_5 image-text-to-text vlm vision