Category
Data agents
7,722 Data AI agents indexed on MeshKore — the most complete public catalog, ranked by popularity and updated daily.
7,722 agents · ranked by popularity · refine in the directory →
Data agents — page 70 of 78
A simple RAG (Retrieval-Augmented Generation) framework for Python.
An up-to-date simple random user-agent with real world database.
Client for the Research Data Storage Registry, Research Computing Services, University of Melbourne.
reshape data type in py language
reeln-cli plugin for OpenAI-powered LLM integration (metadata, translation, zoom)
Research Agent AI Agent Directory to Host All Research Agent related AI Agents Services, Community, Reviews and More.
Research API and SDK
Intelligent research paper analysis pipeline with LLM-driven categorization
A Python package for creating a research brief agent
Automated Research. Powered by LoopGPT.
Aids users in publishing data to a Record Evolution Datapod
A Python library that meters LiteLLM usage to Revenium with context-based metadata injection and framework integrations.
Fast, efficient, minimal, extendible and elegant RAG system
An EOS agent to collect and process RPKI VRP data
RPX — wrap any robot training command with full end-to-end analytics.
Run a Python script with coverage tracking and allow the user to specify the coverage data file.
Shadow-Sandbox DB Layer -- let AI agents modify your database safely with tenant isolation, Pydantic validation, and atomic sync.
A document-ingesting agent that monitors specified directories, keeping stored documents up to date in a vector database for Retrieval-Augmented Generation (RAG) queries.
Sandlake Storage SDK - 基于 S3 存储后端的模型和数据集下载/上传 SDK,提供类似 ModelScope 风格的 API
LLM components for the Sayou Data Platform
A tool for comparing populations in single-cell RNA-seq data with average overlap of marker gene lists
Supply chain compromise scanner — detects known PyPI and npm attacks via data-driven threat profiles
Secure Cloud Data Migration Agent (Linux local agent)
LLM-driven agent for describing data tables based on domain schemas
Multi-agent scientific literature research system with persistent memory
An open-source agentic harness for scientific literature analysis — multi-model data extraction, systematic reviews, and provenance tracking
Python library for scraping ChatGPT
Use a random User-Agent provided by fake-useragent for every request
Use a random User-Agent provided by fake-useragent for every request
A Scrapy extension for data extraction using LangChain
Automatically pick an User-Agent for every request
Purpose-scoped ADK agents for SDC4 data operations
TOON-Native Auto-Embedding & RAG Toolkit for MariaDB — VECTOR(N), HNSW, VEC_DISTANCE_COSINE, with TOON tabular output that saves 10-55% LLM tokens vs JSON.
A taskflow agent for the SecLab project, enabling secure and automated workflow execution.
Tools for extract sensitive configuration out of your project
AI-powered dataset segmentation agent for ML workflows
Add your description here
Semantic Kernel plugins for Built-Simple research APIs (PubMed, ArXiv, Wikipedia)
SGR Agent Core - Schema-Guided Reasoning for building agent
Local package containing the ShareGPT V3 unfiltered cleaned split dataset.
Simple HTTP client for SheetBase - Use Google Sheets as a database
SIA: Self-Improving AI framework
A small, easy to use module that helps you store data easily.
This package will help you talk to your data using Retrieval Augmented Generation (RAG).
Package that provides support for Langchain community data loaders.
SincroX Agent — lightweight SQL bridge that runs next to your database
Orchestrator, Generic Agent, and Research Agent components of the Sirji AI agentic framework.
LLM-driven self-healing API discovery for undocumented SaaS portals via CDP
Bayesian Transfer Learning for Small-Data Predictive Analytics
smappdragon is a set of tools for working with twitter data
Self-aware vector embeddings with temporal awareness, confidence decay, and relational intelligence for RAG systems
A Snakemake storage plugin for the DIRAC Data Management System (DMS).
Data exchange agent for migrations and validation
Research agent using GPT-5.2 to search Reddit and Bluesky
Production-grade Retrieval-Augmented Generation package
Domain-agnostic agent framework for integrating AI agents into data pipelines
Distributed LLM evaluation framework for Apache Spark
AI-powered assistant for spike sorting and neural data analysis
A Python tool for splitting large Markdown files into smaller sections based on a specified token limit. This is particularly useful for processing large Markdown files with GPT models, as it allows the models to handle the data in manageable chunks.
Spectrum instance spotinst-agent that is able to run remote scripts, collect data, deploy applications and more.
Spectrum instance spotinst-agent that is able to run remote scripts, collect data, deploy applications and more.
Spectrum instance spotinst-agent that is able to run remote scripts, collect data, deploy applications and more.
Spectrum instance spotinst-agent that is able to run remote scripts, collect data, deploy applications and more.
Spryngtime Usage Analytics & Billing API
Spryngtime Usage Analytics & Billing API
A secure SQL tool for AI agents to interact with databases
A tool to extract SQL database schema and generate ChatGPT prompts
SQL agent, query your database
Conversational Text-to-SQL agent that allows you to ask natural language questions over structured data (Excel → SQLite), using semantic search, automatic SQL generation, safe execution, and LLM-based result interpretation — all with conversation memory and SQL history.
SQL agent, query your database
Service-first database intelligence platform for safe inspection and agent-ready exports
Python package that loads data from Minecraft: Bedrock Edition packs into SQLite database
A Streamlit component enabling LLM-driven semantic search and data retrieval.
use llm to clean data for start ai
STAT: Spatial Transcriptomics Analytical agenT - AI-powered platform for spatial omics analysis
Statistics and data analysis toolkit
STELLA: Self-Evolving Multimodal Agents for Biomedical Research
Local Stitchflow daemon for browser automation tasks and data sync
A configuration-driven stock data storage framework
Make it easy to work with local or cloud storage as part of a data science workflow
Scan file-system storage, record results in a database, and serve a live web UI.
A CLI tool for creating and submitting SciCat dataset entries
A simple Databricks package
A powerful framework for building AI agents with Claude
A professional-grade AI utility for automated data synchronization and backend management.
Sample Research Agent powered by Strands Agents SDK
MCP server that exposes your Strava athletic data as tools for Claude and other LLMs
stream_llm_parser is a Python library for parsing and processing streaming data with special token handling.
Data structuring with LLM
Structured data extraction from text using LLMs and dynamic model generation
Weather forecast data
Supadata document loader integration for LangChain (Python)
Superduper allows users to work with anthropic API models. The key integration is the integration of high-quality API-hosted LLM services.
Superduper allows users to work with cohere API models.
Superduper allows users to work with self-hosted LLM models via [Llama.cpp](https://github.com/ggerganov/llama.cpp).
Superduper allows users to work with openai API models.
Superduper allows users to work with self-hosted LLM models via [vLLM](https://github.com/vllm-project/vllm).
Verified data deletion and leak detection for RAG systems
Sutra-RAG: The Contextual Retriever library bridging raw data with LLM prompts