Skip to main content
📖 The AI Tool Bible

Best AI tools for rag frameworks

48 tools in the RAG category, filtered to rag frameworks.

All RAG
Pinecone preview image
Pinecone logo

Pinecone

Featured
RAG · Hosted vector DB (not an LLM)
8.8

Managed vector database for production-scale similarity search.

Freemium· Starter: Free · Builder: $20/month flat · Standard: $50/month min. usage · Enterprise: $500/month min. usagemanaged vector DBproduction RAG
LlamaIndex preview image
LlamaIndex logo

LlamaIndex

Featured
RAG · BYO (Claude / GPT / open)
8.7

Data framework for connecting LLMs to your data.

Freemium· Free open-source; LlamaCloud paidRAGdata ingestion
Elasticsearch Vector Search preview image
Elasticsearch Vector Search logo

Elasticsearch Vector Search

RAG · BYO embeddings (OpenAI, Cohere, Hugging Face, Mistral, Bedrock, Vertex, Azure) plus Elastic's built-in ELSER sparse model and E5 dense model
8.7

Hybrid vector + keyword search in the enterprise-grade Elasticsearch engine

Freemium· Resource based pricing: Pay as you go (monthly) or prepaid · Usage based pricing: Pay as you go (monthly) or prepaid · License based pricing: ?RAG chatbot over enterprise docsHybrid semantic + keyword product search
Snowflake Cortex preview image
Snowflake Cortex logo

Snowflake Cortex

RAG · Anthropic Claude, Meta Llama, Mistral Large 2, Snowflake Arctic
8.7

Generative AI and RAG built into the Snowflake data cloud

Enterprise· Standard: Contact sales · Enterprise: Contact sales · Business Critical: Contact sales · Virtual Private Snowflake: Contact salesEnterprise RAG chatbot over governed dataNatural-language SQL for business analysts
DataStax Astra DB preview image
DataStax Astra DB logo

DataStax Astra DB

RAG · Bring-your-own embeddings; integrates with OpenAI, Cohere, Hugging Face, Mistral, NVIDIA NIM, and Vertex AI via server-side vectorize
8.6

Serverless vector and document database for production RAG and AI agents

Freemium· Small On-Demand: Contact sales · Medium (Balanced): Contact sales · Medium (Storage Optimized): Contact sales · Large (Balanced): Contact sales · Large (Storage Optimized): Contact salesRAG chatbot over enterprise documentsAgent long-term memory store
MongoDB Atlas Vector Search preview image
MongoDB Atlas Vector Search logo

MongoDB Atlas Vector Search

RAG · Bring-your-own embeddings (OpenAI, Cohere, open models); native Voyage AI embeddings and rerankers
8.6

Vector search built into the operational database you're already using.

Freemium· Free: $0 · Flex: Up to $30 · Dedicated: Starts at $56.94RAG over enterprise documentsProduct and content recommendation engines
Quivr preview image
Quivr logo

Quivr

RAG · Multi-model (OpenAI, Anthropic, Mistral, Gemma)
8.4

Open-source RAG framework for building custom AI assistants over your own documents in a few lines of Python.

Free· Open source (pip install quivr-core); pay only for LLM/vector-store usagedocument-qacustom-knowledge-base
Weaviate preview image
Weaviate logo

Weaviate

RAG · Hosted vector DB (not an LLM)
8.4

Open-source vector DB with hybrid search and modules.

Freemium· Free: $0 · Flex: $45 · Premium: $400self-hosted RAGhybrid search
LangChain preview image
LangChain logo

LangChain

RAG · BYO (any major LLM)
8.3

The broad LLM application framework — chains, agents, retrievers.

Freemium· Free open-source; LangSmith paidgeneral LLM appsRAG
Vanna.ai preview image
Vanna.ai logo

Vanna.ai

RAG · Multi-model (Anthropic, OpenAI, Gemini, Ollama)
8.3

Open-source text-to-SQL agent that learns your schema and writes queries against your real warehouse.

Freemium· Explorer: $50 · Team: $500 · Enterprise: Contact salestext-to-sqlnatural-language-bi
Feast preview image
Feast logo

Feast

RAG
8.2

Open-source feature store that serves consistent features to ML training and online inference, with RAG vector search built in.

Free· Free, open source (Apache 2.0); self-hostedfeature-storerag-retrieval
LanceDB preview image
LanceDB logo

LanceDB

RAG
8.2

Open-source multimodal lakehouse and vector database built for AI training and retrieval at petabyte scale.

Freemium· Open-source free; LanceDB Cloud and Enterprise via contact salesvector-searchrag
Chroma preview image
Chroma logo

Chroma

RAG · Hosted vector DB (not an LLM)
8.1

Embedded, developer-friendly vector store for Python.

Freemium· Starter: $0 · Team: $250 · Enterprise: Customprototypingembedded RAG
Databricks Vector Search preview image
Databricks Vector Search logo

Databricks Vector Search

RAG · Multi-model (BYO embeddings or Databricks-hosted)
8.1

Managed hybrid vector search that lives inside the Databricks lakehouse and auto-syncs with your source tables.

Enterprise· Standard: $605 · Storage Optimized: $922rag-retrievalhybrid-search
RAGFlow preview image
RAGFlow logo

RAGFlow

RAG · Multi-model
8.1

Open-source RAG engine with deep document parsing, hybrid search, and visual agent orchestration.

Freemium· Free tier; Starter $29/mo; Pro $129/mo; Enterprise customdocument-qaenterprise-search
Exa preview image
Exa logo

Exa

RAG · Proprietary neural + keyword search
8.0

Web search API built for AI agents, with structured outputs and token-efficient highlights.

Freemium· Free Tier: Free · Search: $7/1k requests · Agent: $0.012–$1.00/run · Contents: $1/1k pages per content type · Deep Search: $12–15/1k requestsagent-web-searchrag-retrieval
Firecrawl preview image
Firecrawl logo

Firecrawl

RAG · Claude, Cursor, Windsurf, OpenAI, Gemini
8.0

Web scraping and crawling API that returns LLM-ready markdown, JSON, or structured data from any URL.

Freemium· Free Plan: $0 · Hobby: $16 · Standard: $83 · Growth: $333 · Scale: $599/monthlyweb-scrapingrag-ingestion
AnythingLLM preview image
AnythingLLM logo

AnythingLLM

RAG · Multi-model
7.9

Open-source desktop and self-hosted app that turns your documents into a private chat-and-agent workspace.

Freemium· Basic: $50/monthly · Pro: $99/monthly · Enterprise: Contact Usdocument-chatprivate-rag
Langchain-Chatchat preview image
Langchain-Chatchat logo

Langchain-Chatchat

RAG · Multi-model (GLM-4, Qwen2, Llama 3, etc. via Xinference/Ollama/LocalAI/FastChat)
7.4

Self-hostable RAG and agent framework that wires LangChain to any local open-source LLM and a knowledge base.

Free· Apache-2.0 open source; self-hosted, infra costs onlyprivate-knowledge-baseoffline-rag
Agentset preview image
Agentset logo

Agentset

RAG · Multi-model (Claude, OpenAI, Google, xAI, Cohere, Mistral, DeepSeek)
7.3

Production-ready RAG infrastructure with agentic search, citations, and model-agnostic plumbing.

Freemium· Free: $0 · Pro: $49 · Enterprise: Customdocument-qaagentic-search
Epsilla preview image
Epsilla logo

Epsilla

RAG · Multi-model
7.3

Agent-as-a-Service platform with managed RAG and a no-code builder for vertical enterprise AI.

Freemium· Free Tier: $0/month · Starter Tier: $29/month · Professional Tier: $249/month · AI Concierge: $2,499/month · Enterprise Tier: Custom/monthenterprise-ragai-agents
Pathway preview image
Pathway logo

Pathway

RAG · Multi-model
7.3

Live data framework for production RAG and streaming ETL pipelines in Python.

Freemium· Community free (BSL 1.1, 8GB/4 cores); Scale and Enterprise tiers with license keylive-ragstreaming-etl
Cognee preview image
Cognee logo

Cognee

RAG · Multi-model (Claude, OpenAI, others)
7.2

Open-source graph-memory layer that gives AI agents persistent, queryable context across sessions.

Freemium· Hobby free (1M tokens/mo); Growth $5/workspace/mo + token usage; Enterprise customagent-memoryknowledge-graphs
Perplexity AI preview image
Perplexity AI logo

Perplexity AI

RAG · Multi-model (Sonar, GPT-4 class, Claude, Gemini)
7.2

Conversational answer engine that cites its sources by default.

Freemium· Basic: $10 · Pro: $30 · Enterprise: Contact salesai-searchresearch
BGE (BAAI General Embedding) preview image
BGE (BAAI General Embedding) logo

BGE (BAAI General Embedding)

RAG · BGE / bge-m3 / bge-reranker
7.1

Open-source embedding and reranker models from BAAI that anchor a huge share of production RAG stacks.

Free· Free, open-source (MIT-style license); self-hosted inference cost onlysemantic-searchrag-retrieval
OpenDataLoader PDF preview image
OpenDataLoader PDF logo

OpenDataLoader PDF

RAG
7.1

Open-source PDF parser built for RAG pipelines, with reading-order detection, table extraction, and bounding-box citations.

Freemium· Free (Apache 2.0); enterprise tier for PDF/UA export and visual editorpdf-parsingrag-preprocessing
PostgresML preview image
PostgresML logo

PostgresML

RAG · Multi-model (Llama, Mistral, open-source embeddings)
7.1

PostgreSQL extension that runs embeddings, vector search, and LLM inference inside your database.

Freemium· Serverless: From $7.50 per query hour · Dedicated: From $0.60 per instance hour · Enterprise: Custom pricingvector-searchrag
Rivestack preview image
Rivestack logo

Rivestack

RAG · OpenAI embeddings (auto-embeddings)
7.1

Managed Postgres with pgvector on dedicated NVMe, pitched as a cheaper RAG backend than Pinecone or Supabase.

Freemium· Free: $0/month · Solo: $29/month · HA Cluster: Starting at $49/node/monthrag-backendvector-search
UltraRAG preview image
UltraRAG logo

UltraRAG

RAG · Multi-model (MiniCPM-Embedding-Light, AgentCPM-Report, BYO LLM)
7.1

Low-code, YAML-driven RAG pipeline orchestrator with a visual UI for building and demoing retrieval systems.

Free· Open source; self-hostedrag-pipelinesknowledge-base-qa
GaliChat preview image
GaliChat logo

GaliChat

RAG
7.0

No-code AI chatbot builder that trains on your website content for support and lead capture.

Freemium· pro: $19 · business: $99 · enterprise: $999 · Free Demo: Freecustomer-supportlead-generation
HelixDB preview image
HelixDB logo

HelixDB

RAG
7.0

Unified graph-and-vector database built for AI agent memory and GraphRAG.

Freemium· GW-10: $86.87 · GW-20: $173.74 · GW-40: $348.21 · GW-80: $696.42 · GW-160: $1,392.84agent-memorygraphrag
Kotaemon preview image
Kotaemon logo

Kotaemon

RAG · Multi-model (OpenAI, LlamaCPP, any OpenAI-compatible endpoint)
7.0

Open-source RAG UI for chatting with your own documents, locally or self-hosted.

Free· Free, open-source (MIT-style); self-hosted infrastructure costs onlydocument-qaprivate-rag
MaxKB preview image
MaxKB logo

MaxKB

RAG · Multi-model
7.0

Open-source enterprise RAG and agent platform with built-in workflow engine and multi-LLM support.

Freemium· Community edition free (GPLv3); paid enterprise editionenterprise-knowledge-basecustomer-support-bots
PageIndex preview image
PageIndex logo

PageIndex

RAG
7.0

Vectorless reasoning-based retrieval for long documents, with traceable, auditable answers.

Freemium· Free Try Now tier; enterprise pricing on requestdocument-qalong-pdf-retrieval
PrivateGPT preview image
PrivateGPT logo

PrivateGPT

RAG · Multi-model (BYO local LLM)
7.0

Production-ready, air-gapped RAG framework for querying your documents with local LLMs.

Freemium· OSS free; Zylon enterprise contract (contact sales)private-ragchat-with-documents
RAGs by LlamaIndex preview image
RAGs by LlamaIndex logo

RAGs by LlamaIndex

RAG · Multi-model (OpenAI, Anthropic, Replicate, HuggingFace)
7.0

Open-source Streamlit app that builds a custom RAG pipeline from a natural-language brief.

Free· Free, MIT-licensed; bring your own model/API keysnatural-language-rag-builderdocument-qa
Singlebase Cloud preview image
Singlebase Cloud logo

Singlebase Cloud

RAG · Multi-model
7.0

AI-native Firebase alternative bundling document DB, vector DB, auth, storage, and built-in AI services.

Freemium· Free tier available; paid plans scale with usagevector-searchrag-apps
Superduper preview image
Superduper logo

Superduper

RAG · Multi-model
7.0

Enterprise AI agent orchestration that brings RAG and agents to your existing data stack without migration.

Enterprise· Free trial on Snowflake Marketplace; enterprise self-hosted pricing on requestin-database-ragagent-orchestration
CocoIndex preview image
CocoIndex logo

CocoIndex

RAG · Bring-your-own (embeddings + LLM)
6.9

Open-source incremental data framework that keeps RAG indexes and agent context continuously fresh.

Free· Open-source, self-hosted; bring your own infracode-indexingrag-pipelines
Cohere preview image
Cohere logo

Cohere

RAG · Command, Embed, Rerank, Transcribe (proprietary)
6.9

Enterprise-grade LLM platform built for private, secure, and customizable deployment.

Enterprise· Embed 4 Small: $2,500 · Embed 4 Medium: $3,250 · Rerank 3.5 Medium: $3,250 · Rerank 4 Fast Medium: $3,250 · Rerank 4 Pro Medium: $3,250enterprise-ragsemantic-search
DeepSearcher preview image
DeepSearcher logo

DeepSearcher

RAG · Multi-model (DeepSeek, OpenAI o1/o3-mini, Claude, Llama, others)
6.9

Open-source agentic RAG framework for private enterprise data, built by the Zilliz/Milvus team.

Free· Free, Apache 2.0; bring your own LLM and vector DB costsenterprise-ragagentic-search
Haystack preview image
Haystack logo

Haystack

RAG · Multi-model
6.9

Open-source Python framework from deepset for building production RAG pipelines and LLM agents.

Freemium· Open-source free; deepset Enterprise Support and AI Platform via salesragagents
You.com preview image
You.com logo

You.com

RAG · Multi-model
6.9

Web search and research APIs purpose-built for LLMs and AI agents.

Freemium· Free trial; enterprise pricing on requestweb-search-apiagent-grounding
Context Data preview image
Context Data logo

Context Data

RAG · Multi-model
6.8

Enterprise data platform for deploying private RAG pipelines without infrastructure plumbing.

Enterprise· Contact salesenterprise-ragdocument-search
TurboVec preview image
TurboVec logo

TurboVec

RAG
6.8

Rust-powered vector index with 2-4 bit TurboQuant compression for SIMD-accelerated RAG search.

Free· Free, MIT licensedvector-searchrag
Yuxi preview image
Yuxi logo

Yuxi

RAG · Multi-model
6.8

Open-source AI agent platform that fuses agentic RAG with knowledge graphs on a LangGraph runtime.

Free· Free, MIT-licensed self-hostagentic-ragknowledge-graphs
ClickHouse preview image
ClickHouse logo

ClickHouse

RAG

The open-source columnar database powering real-time analytics — and, increasingly, LLM observability and RAG backends.

Freemium· Open-source self-managed: free. ClickHouse Cloud: from $50/month (usage-based on compute + storage, AWS/GCP/Azure). Enterprise tier available with dedicated support and BYOC options.LLM trace and cost analyticsRAG retrieval with hybrid vector + metadata filters
Docling preview image
Docling logo

Docling

RAG · GraniteDocling 258M and other vision-language models

Open-source document parsing for AI: PDFs, Office files, audio and video into clean, structured Markdown/JSON

Free· Free and open source under MIT License. Hosted by the LF AI & Data Foundation; no paid tiers.RAG document ingestionPDF table extraction