Category
Data agents
7,722 Data AI agents indexed on MeshKore — the most complete public catalog, ranked by popularity and updated daily.
7,722 agents · ranked by popularity · refine in the directory →
Data agents — page 66 of 78
A library of community-driven data loaders for LLMs. Use with LlamaIndex and/or LangChain.
Llama-index integration with Desearch API for search and data-fetching tools.
llama-index embeddings databricks integration
llama-index embeddings OCI Data Science integration
Interface between LLMs and your data
LlamaIndex Graph Stores integration for ArcadeDB - Multi-Model Database with Graph, Document, Key-Value, Vector, and Time-Series support
llama-index llms databricks integration
llama-index llms OCI Data Science integration
llama-index LLM integration with You.com's conversational Smart and Research APIs
A dataset generator for llama index to aid in generating ORPO & DPO datasets
llama-index packs diff private simple dataset
llama-index packs gmail_openai_agent integration
llama-index packs llama_dataset_metadata integration
llama-index packs RAFT Dataset paper implementation
llama-index packs stock_market_data_query_engine integration
llama-index packs tables integration
A Salesforce NPSP data reader for LlamaIndex to empower data-driven fundraising intelligence.
llama-index readers apify integration
llama-index readers athena integration
llama-index readers bagel integration
LlamaIndex readers for Built-Simple research APIs (PubMed, ArXiv, Wikipedia)
llama-index readers database integration
llama-index readers HuggingFace Datasets integration
llama-index readers firebase_realtimedb integration
llama-index readers firestore integration
llama-index readers pdb integration
llama-index readers scrappey integration
llama-index readers semanticscholar integration
llama-index readers snowflake integration
ServiceNow data reader for LlamaIndex - sync and async readers for Incidents, CMDB, KB, Changes, Problems, Catalog, and Attachments.
llama-index readers structured_data integration
llama-index readers wordlift integration
llama-index readers youtube-metadata integration
llama-index tools arxiv integration
llama-index tools Bright Data integration
llama-index tools database integration
llama-index tools integrating DuckDuckGo search
llama-index tools linkup_research integration
LlamaIndex tool wrappers for MAXIA Oracle — multi-source price feeds for AI agents. Data feed only.
llama-index tools integrating ScrapegraphAI
llama-index tools to use ScraperAPI web scraping
llama-index tools tavily_research integration
AI framework integrations for Azure Database for PostgreSQL
llama-index vector_stores databricks vector search integration
llama-index vector_stores db2 database integration
Vector Database for Fast ANN Searches
LlamaIndex integration for HyperspaceDB - Hyperbolic Vector Database
llama-index vector_stores oracle database integration
LlamaIndex VectorStore integration for Pixeltable multimodal data infrastructure.
LlamaIndex VectorStore for VelesDB: The Local AI Memory Database. Microsecond RAG retrieval.
ZeroDB vector store for LlamaIndex - AI-native vector database with free embeddings, semantic search, and RAG support. Pinecone alternative.
LlamaIndex integration for ZeusDB vector database. Enterprise-grade RAG with high-performance vector search.
Easy deployment of quantized llama models on cpu
Data processing pipeline using MLX (scraper, chunker, extractor).
Data processing pipeline using MLX (scraper, chunker, extractor).
A comprehensive simulation framework for AI research and testing scenarios
Web scraping tool potentially using Llama models.
Blockchain intelligence and analytics platform
Dataset management and processing library for LlamaSearch.ai applications
Next-Gen Hybrid Python/Rust Data Platform with MLX
Database management and query optimization library for Python
The soul ecosystem for LlamaIndex: persistent memory, identity, database schema intelligence, and SoulMate API integration.
A powerful library for AI-powered search and data processing
A comprehensive PDF processing toolkit for document workflows
LLAMASS is a Loader for the AMASS dataset
An easy-to-extend LLM annotator for robust, resumable data annotation.
LLM-A*: Large Language Model Enhanced Incremental Heuristic Search on Path Planning.
An intelligent automated exploratory data analysis tool powered by Large Language Models, providing in-depth insights and visualizations for your datasets
Batch-oriented LLM annotation workflows for tabular datasets with OpenAI Batch support.
A package to process CSV text data in batches using OpenAI API
Evaluate large-language models for undesirable behaviors such as bias.
Benchmark LLMs with 10 benchmarks & 132K+ questions. 8 providers: OpenAI, Anthropic, Groq, Together, Fireworks, DeepSeek, Ollama, HuggingFace. Unified CLI + Web dashboard.
Semantic-aware chunking with provenance tracking for production RAG and LLM data pipelines
Real-time cost tracking, budget enforcement, and usage analytics for LLM applications
Best open-source document to markdown converter for LLM training data. Convert PDF, Word, PowerPoint, Excel, images, URLs to clean markdown, JSON, HTML locally. Alternative to Unstructured, Docling, Marker, MarkItDown, MinerU, PaddleOCR, Tesseract
A Python package for calculating key metrics to assess LLM performance in various tasks, including extracting structured dataa.
A processor for LLM tasks
Generate synthetic evaluation datasets from your own documents.
LLM access to Databricks model serving
A dataclass interface for llms
极简高性能流式数据加工库
Python3 library for converting between various LLM dataset formats.
Meta-library that combines all llm-dataset-converter libraries.
A collection of datasets for language model training including scripts for downloading, preprocesssing, and sampling.
RAG llm datatech
LLM plugin of various deep research implementations
Intelligent LLM dispatching with performance-based routing, multimodal support, streaming, monitoring, and comprehensive analytics
Advanced Knowledge Graph Engine with Document Processing, Semantic Search and Multi-LLM Integration
Tools that use LLM to explain datasets
A Lakehouse LLM Explorer. Wrapper for spark, databricks and langchain processes
A framework that enables efficient extraction of structured data from unstructured text using large language models (LLMs).
Extract structured, validated JSON from any LLM  OpenAI, Anthropic, Gemini  with batch extraction, caching, per-field confidence scoring, schema evolution, multi-schema extraction, output transforms, partial extraction, extraction diff, pipeline extraction, and smart auto-retry.
Automated feature engineering using Large Language Models (LLMs) for tabular data
Generate interpretable feature schemas and tabular exports from text, images, tabular data, and video with LLMs
Token-efficient data format converter for LLM contexts
LLM fragments plugin for PyPI packages metadata
Minimal metadata and discovery helpers for typed Python tools used by llm_function runtimes.
LLM-Guard is a comprehensive tool designed to fortify the security of Large Language Models (LLMs). By offering sanitization, detection of harmful language, prevention of data leakage, and resistance against prompt injection attacks, LLM-Guard ensures that your interactions with LLMs remain safe and secure.
LLM-optimized HTML cleaning: hydration extraction, token budgets, multiple output formats
a data preprocessing toolkit that makes it easy to create common LLM-related data structures; from training data to chain payloads!