Category
Data agents
6,796 Data AI agents indexed on MeshKore — the most complete public catalog, ranked by popularity and updated daily.
6,796 agents · ranked by popularity · refine in the directory →
Data agents — page 63 of 68
Self-aware vector embeddings with temporal awareness, confidence decay, and relational intelligence for RAG systems
A Snakemake storage plugin for the DIRAC Data Management System (DMS).
Data exchange agent for migrations and validation
Research agent using GPT-5.2 to search Reddit and Bluesky
Production-grade Retrieval-Augmented Generation package
Domain-agnostic agent framework for integrating AI agents into data pipelines
Distributed LLM evaluation framework for Apache Spark
AI-powered assistant for spike sorting and neural data analysis
A Python tool for splitting large Markdown files into smaller sections based on a specified token limit. This is particularly useful for processing large Markdown files with GPT models, as it allows the models to handle the data in manageable chunks.
Spectrum instance spotinst-agent that is able to run remote scripts, collect data, deploy applications and more.
Spectrum instance spotinst-agent that is able to run remote scripts, collect data, deploy applications and more.
Spectrum instance spotinst-agent that is able to run remote scripts, collect data, deploy applications and more.
Spectrum instance spotinst-agent that is able to run remote scripts, collect data, deploy applications and more.
Spryngtime Usage Analytics & Billing API
Spryngtime Usage Analytics & Billing API
A secure SQL tool for AI agents to interact with databases
A tool to extract SQL database schema and generate ChatGPT prompts
SQL agent, query your database
Conversational Text-to-SQL agent that allows you to ask natural language questions over structured data (Excel → SQLite), using semantic search, automatic SQL generation, safe execution, and LLM-based result interpretation — all with conversation memory and SQL history.
SQL agent, query your database
Service-first database intelligence platform for safe inspection and agent-ready exports
Python package that loads data from Minecraft: Bedrock Edition packs into SQLite database
A Streamlit component enabling LLM-driven semantic search and data retrieval.
use llm to clean data for start ai
STAT: Spatial Transcriptomics Analytical agenT - AI-powered platform for spatial omics analysis
Statistics and data analysis toolkit
STELLA: Self-Evolving Multimodal Agents for Biomedical Research
Local Stitchflow daemon for browser automation tasks and data sync
A configuration-driven stock data storage framework
Make it easy to work with local or cloud storage as part of a data science workflow
Scan file-system storage, record results in a database, and serve a live web UI.
A CLI tool for creating and submitting SciCat dataset entries
A simple Databricks package
A powerful framework for building AI agents with Claude
A professional-grade AI utility for automated data synchronization and backend management.
Sample Research Agent powered by Strands Agents SDK
MCP server that exposes your Strava athletic data as tools for Claude and other LLMs
stream_llm_parser is a Python library for parsing and processing streaming data with special token handling.
Data structuring with LLM
Structured data extraction from text using LLMs and dynamic model generation
Weather forecast data
Supadata document loader integration for LangChain (Python)
Superduper allows users to work with anthropic API models. The key integration is the integration of high-quality API-hosted LLM services.
Superduper allows users to work with cohere API models.
Superduper allows users to work with self-hosted LLM models via [Llama.cpp](https://github.com/ggerganov/llama.cpp).
Superduper allows users to work with openai API models.
Superduper allows users to work with self-hosted LLM models via [vLLM](https://github.com/vllm-project/vllm).
Verified data deletion and leak detection for RAG systems
Sutra-RAG: The Contextual Retriever library bridging raw data with LLM prompts
A small crawler to scrape data from swranking.com and store in in the local db of the vm.
Generate realistic API data from Swagger json using Groq
MCP server for Foursquare Swarm check-in data
Automated research paper tracking and knowledge synthesis
Create a SQLite database containing your checkin history from Foursquare Swarm
PEP 458 compatible detached signing provider for Swarmauri
Web Scraping Tool for Swarmauri
CLI for the Swarm & Bee dataset bakery — order curated AI training corpora from the terminal.
Calculate Field Aligned Currents based on swarm data, mainly through the ViRES python interface viresclient. Some utilities for these kinds of scientific calculations are also provided.
Add your description here
LLM-driven agent for categorizing data tables based on domain schemas
Tabsdata Agent service for AI-powered data processing.
Seamless Feature Extraction and Interpretation of Text Columns in Tabular Data Using Large Language Models
Tair vector database with Haystack
Singer.io tap for extracting data from Yahoo Gemini
Singer.io target for extracting data
A collection of tools and agents for building AI applications with Tavily
Export teamcity agent data
Tencent Cloud Dataagent SDK for Python
Create EO Minicubes from Polygons and simplify EO Data downloading.
A flexible, LLM-agnostic text-to-SQL agent that works with any database
Module used to interact with Terragrunt and OpenTofu/Terraform state data.
check for different services e.g: Redis, SSL sertificate, Sites, Elasticsearch,Database
Storage and database adapters available in project Thoth
Storage and database adapters available in project Thoth
A10 Thunder Observability Agent is a lightweight autonomous data processing engine that can be externally installed and configured for multiple A10 Thunder Instances to collect, process and publish performance metrics and logs into multiple monitoring dashboard.
This package contains tools to work with Tilores entity resolution database within Langchain.
Tiny library for key-value single-file application data storage
Tiny library for key-value single-file application data storage
Tiny library for key-value single-file application data storage
A Python library for performing deep research using AI agents and Firecrawl, with support for custom models and report generation.
Given impression and KPI data, determine the tipping points associated with maximizing success rate of marketing
A python module for TM1 Bedrock.
Tools for LLM
A collection of database and RAG operation tools
A collection of database and RAG operation tools
Toroidal topology primitives for LLM coherence research (v7: replication update — inference-time bias null result, prompt hardening effective)
A topological data analysis library for detecting knowledge gaps in RAG systems.
Comprehensive TPS monitoring and throttling library with database-driven controls
TrainLoop LLM Logging SDK for data collection
A library for transforming unstructured text into structured data without context/mappings using ChatGPT.
A Python library for accessing Google Trends data
A Python library for accessing Google Trends data
Free, open-source Python library for Google Trends data: trending now plus keyword interest over time, related queries, and interest by region. A maintained pytrends alternative with CLI and API.
Claim-grounded truth evaluation for LLM outputs: retrieval, per-claim verification, and explicit risk signals
多论文 RAG 问答系统 - 基于检索增强生成的科研论文问答系统
SQLite schema, Python classes and data management tools for working with time series from engineering test beds
Compressed vector and graph-augmented retrieval engine with adaptive two-stage search, implementing TurboQuant/QJL quantization from Google Research.
A package which handles the storage of TVB data
Python interface to the TVRage television information database.
Panel-based reliability annotation for preference data: measure inter-judge agreement (IJA) with a panel of LLM judges and export TRL-ready DPO data with reliability signals attached.