Skip to main content
📖 The AI Tool Bible

The RAG stack that actually scales

Twelve pieces — embedding, vector store, retrieval, eval, framework — that hold up past a prototype.

The RAG hello-world looks easy. The production RAG stack is a different beast: you need an embedding model, a vector store, a retrieval framework, a reranker, and an eval harness that catches regressions before users do. These twelve tools cover every layer. Pick one from each — don't pick two from the same.

13 tools in this collection
Pinecone preview image
Pinecone logo

Pinecone

Featured
RAG · Hosted vector DB (not an LLM)
8.8

Managed vector database for production-scale similarity search.

Freemium· Starter: Free · Builder: $20/month flat · Standard: $50/month min. usage · Enterprise: $500/month min. usagemanaged vector DBproduction RAG
Weaviate preview image
Weaviate logo

Weaviate

RAG · Hosted vector DB (not an LLM)
8.4

Open-source vector DB with hybrid search and modules.

Freemium· Free: $0 · Flex: $45 · Premium: $400self-hosted RAGhybrid search
Chroma preview image
Chroma logo

Chroma

RAG · Hosted vector DB (not an LLM)
8.1

Embedded, developer-friendly vector store for Python.

Freemium· Starter: $0 · Team: $250 · Enterprise: Customprototypingembedded RAG
Vespa preview image
Vespa logo

Vespa

RAG · Hosted search engine (not an LLM)
8.2

Yahoo's open-source search engine with vector + sparse retrieval.

Freemium· Free open-source; Vespa Cloud paidlarge-scale searchranking
LlamaIndex preview image
LlamaIndex logo

LlamaIndex

Featured
RAG · BYO (Claude / GPT / open)
8.7

Data framework for connecting LLMs to your data.

Freemium· Free open-source; LlamaCloud paidRAGdata ingestion
LangChain preview image
LangChain logo

LangChain

RAG · BYO (any major LLM)
8.3

The broad LLM application framework — chains, agents, retrievers.

Freemium· Free open-source; LangSmith paidgeneral LLM appsRAG
Feast preview image
Feast logo

Feast

RAG
8.2

Open-source feature store that serves consistent features to ML training and online inference, with RAG vector search built in.

Free· Free, open source (Apache 2.0); self-hostedfeature-storerag-retrieval
RAGFlow preview image
RAGFlow logo

RAGFlow

RAG · Multi-model
8.1

Open-source RAG engine with deep document parsing, hybrid search, and visual agent orchestration.

Freemium· Free tier; Starter $29/mo; Pro $129/mo; Enterprise customdocument-qaenterprise-search
Humata.ai preview image
Humata.ai logo

Humata.ai

RAG · Multi-model
7.8

Chat-with-your-documents RAG tool with citation-backed answers across uploaded PDFs and files.

Freemium· Free: $0 · Expert: $9.99 · Team: $49 / user · Enterprise: customdocument-qaresearch-summarization
Exa preview image
Exa logo

Exa

RAG · Proprietary neural + keyword search
8.0

Web search API built for AI agents, with structured outputs and token-efficient highlights.

Freemium· Free Tier: Free · Search: $7/1k requests · Agent: $0.012–$1.00/run · Contents: $1/1k pages per content type · Deep Search: $12–15/1k requestsagent-web-searchrag-retrieval
NotebookLM preview image
NotebookLM logo

NotebookLM

RAG · Gemini 2.5
8.1

Google's source-grounded research notebook that turns your documents into chats, briefs, and AI-hosted podcasts.

Freemium· Free tier; Plus via Google One AI Premium ($19.99/mo) or Workspace add-ondocument Q&Aresearch synthesis
Scite preview image
Scite logo

Scite

RAG · Multi-model
8.2

AI research assistant that grades citations as supporting, contrasting, or mentioning across 1.6B citation statements.

Freemium· Basic: $20 · Pro: $50 · Team: $50 · Enterprise: Contact usliterature-reviewcitation-analysis
Cohere preview image
Cohere logo

Cohere

RAG · Command, Embed, Rerank, Transcribe (proprietary)
6.9

Enterprise-grade LLM platform built for private, secure, and customizable deployment.

Enterprise· Embed 4 Small: $2,500 · Embed 4 Medium: $3,250 · Rerank 3.5 Medium: $3,250 · Rerank 4 Fast Medium: $3,250 · Rerank 4 Pro Medium: $3,250enterprise-ragsemantic-search