Best AI tools for retrievers
48 tools in the RAG category, filtered to retrievers.
Pinecone
FeaturedManaged vector database for production-scale similarity search.
LlamaIndex
FeaturedData framework for connecting LLMs to your data.
Elasticsearch Vector Search
Hybrid vector + keyword search in the enterprise-grade Elasticsearch engine
Quivr
Open-source RAG framework for building custom AI assistants over your own documents in a few lines of Python.
Weaviate
Open-source vector DB with hybrid search and modules.
LangChain
The broad LLM application framework — chains, agents, retrievers.
Vanna.ai
Open-source text-to-SQL agent that learns your schema and writes queries against your real warehouse.
Feast
Open-source feature store that serves consistent features to ML training and online inference, with RAG vector search built in.
LanceDB
Open-source multimodal lakehouse and vector database built for AI training and retrieval at petabyte scale.
Scite
AI research assistant that grades citations as supporting, contrasting, or mentioning across 1.6B citation statements.
Vespa
Yahoo's open-source search engine with vector + sparse retrieval.
Chroma
Embedded, developer-friendly vector store for Python.
Cube
Semantic layer that grounds LLM agents in your real business metrics instead of letting them hallucinate SQL.
Databricks Vector Search
Managed hybrid vector search that lives inside the Databricks lakehouse and auto-syncs with your source tables.
NotebookLM
Google's source-grounded research notebook that turns your documents into chats, briefs, and AI-hosted podcasts.
RAGFlow
Open-source RAG engine with deep document parsing, hybrid search, and visual agent orchestration.
Exa
Web search API built for AI agents, with structured outputs and token-efficient highlights.
Firecrawl
Web scraping and crawling API that returns LLM-ready markdown, JSON, or structured data from any URL.
Wren AI
Open-source GenBI semantic layer that lets AI agents query your warehouse in natural language with governed, accurate SQL.
AnythingLLM
Open-source desktop and self-hosted app that turns your documents into a private chat-and-agent workspace.
Humata.ai
Chat-with-your-documents RAG tool with citation-backed answers across uploaded PDFs and files.
Langchain-Chatchat
Self-hostable RAG and agent framework that wires LangChain to any local open-source LLM and a knowledge base.
Agentset
Production-ready RAG infrastructure with agentic search, citations, and model-agnostic plumbing.
Epsilla
Agent-as-a-Service platform with managed RAG and a no-code builder for vertical enterprise AI.
Pathway
Live data framework for production RAG and streaming ETL pipelines in Python.
Rayyan
AI-assisted systematic review platform for screening, deduplicating, and extracting data from large literature corpora.
Cognee
Open-source graph-memory layer that gives AI agents persistent, queryable context across sessions.
Emergent Mind
AI-curated arXiv discovery layer that summarizes frontier papers and aggregates social discussion around them.
FinChat (Fiscal.ai)
AI copilot for equity research that reads filings, transcripts, and KPI tables across 100,000+ public companies.
Graphiti
Open-source temporal knowledge graph framework for building agent memory that updates in real time.
OneKE
Open-source multi-agent framework for schema-guided knowledge extraction from documents.
Perplexity AI
Conversational answer engine that cites its sources by default.
WeKnora
Tencent's open-source RAG framework that turns raw documents into a queryable knowledge base, ReAct agent, and self-maintaining wiki.
BGE (BAAI General Embedding)
Open-source embedding and reranker models from BAAI that anchor a huge share of production RAG stacks.
Elicit
AI research assistant that searches, screens, and extracts data from 138M+ academic papers at scale.
FutureHouse Platform
Multi-agent AI research stack for scientists, with retrieval over 175M+ papers, patents, and trials.
LangExtract
Google's open-source Python library for LLM-driven structured extraction from unstructured text, with source-grounded outputs.
OpenDataLoader PDF
Open-source PDF parser built for RAG pipelines, with reading-order detection, table extraction, and bounding-box citations.
UltraRAG
Low-code, YAML-driven RAG pipeline orchestrator with a visual UI for building and demoing retrieval systems.
alphaXiv
AI reading layer over arXiv with grounded Q&A, auto-summaries, and line-by-line discussion on every preprint.
GaliChat
No-code AI chatbot builder that trains on your website content for support and lead capture.
Graphify
Open-source on-device knowledge graph engine that turns code, docs, papers, meetings and images into a queryable graph.
Kotaemon
Open-source RAG UI for chatting with your own documents, locally or self-hosted.
MaxKB
Open-source enterprise RAG and agent platform with built-in workflow engine and multi-LLM support.
PageIndex
Vectorless reasoning-based retrieval for long documents, with traceable, auditable answers.
PrivateGPT
Production-ready, air-gapped RAG framework for querying your documents with local LLMs.
RAGs by LlamaIndex
Open-source Streamlit app that builds a custom RAG pipeline from a natural-language brief.
Superduper
Enterprise AI agent orchestration that brings RAG and agents to your existing data stack without migration.