Skip to main content
📖 The AI Tool Bible

RAG

Retrieval-augmented generation, vector stores, indexers.

95 tools

Why it matters

RAG isn't a model, it's an architecture — retrieve, augment, generate. The choice is between frameworks that orchestrate the retrieval and the vector stores underneath.

What's in here

Includes RAG frameworks (LlamaIndex, LangChain), managed vector databases (Pinecone), open-source vector stores (Weaviate, Chroma, Vespa), and hybrid-search engines.

How to pick

Pick LlamaIndex when retrieval quality is the bottleneck. Pick Pinecone for zero-ops production. Pick Weaviate or Chroma for self-hosted or budget-conscious. Pick Vespa at scale beyond a few million docs.

Langchain-Chatchat preview image
Langchain-Chatchat logo

Langchain-Chatchat

RAG · Multi-model (GLM-4, Qwen2, Llama 3, etc. via Xinference/Ollama/LocalAI/FastChat)
7.4

Self-hostable RAG and agent framework that wires LangChain to any local open-source LLM and a knowledge base.

Free· Apache-2.0 open source; self-hosted, infra costs onlyprivate-knowledge-baseoffline-rag
Agentset preview image
Agentset logo

Agentset

RAG · Multi-model (Claude, OpenAI, Google, xAI, Cohere, Mistral, DeepSeek)
7.3

Production-ready RAG infrastructure with agentic search, citations, and model-agnostic plumbing.

Freemium· Free: $0 · Pro: $49 · Enterprise: Customdocument-qaagentic-search
Epsilla preview image
Epsilla logo

Epsilla

RAG · Multi-model
7.3

Agent-as-a-Service platform with managed RAG and a no-code builder for vertical enterprise AI.

Freemium· Free Tier: $0/month · Starter Tier: $29/month · Professional Tier: $249/month · AI Concierge: $2,499/month · Enterprise Tier: Custom/monthenterprise-ragai-agents
Pathway preview image
Pathway logo

Pathway

RAG · Multi-model
7.3

Live data framework for production RAG and streaming ETL pipelines in Python.

Freemium· Community free (BSL 1.1, 8GB/4 cores); Scale and Enterprise tiers with license keylive-ragstreaming-etl
Rayyan preview image
Rayyan logo

Rayyan

RAG
7.3

AI-assisted systematic review platform for screening, deduplicating, and extracting data from large literature corpora.

Freemium· Free forever for basic use; enterprise custom pricingsystematic-reviewliterature-screening
Cognee preview image
Cognee logo

Cognee

RAG · Multi-model (Claude, OpenAI, others)
7.2

Open-source graph-memory layer that gives AI agents persistent, queryable context across sessions.

Freemium· Hobby free (1M tokens/mo); Growth $5/workspace/mo + token usage; Enterprise customagent-memoryknowledge-graphs
Emergent Mind preview image
Emergent Mind logo

Emergent Mind

RAG · Undisclosed
7.2

AI-curated arXiv discovery layer that summarizes frontier papers and aggregates social discussion around them.

Freemium· Basic: $0 · Pro: $12 /month · Max: $30 /montharxiv-discoverypaper-summarization
FinChat (Fiscal.ai) preview image
FinChat (Fiscal.ai) logo

FinChat (Fiscal.ai)

RAG · Multi-model (proprietary finance-tuned copilot)
7.2

AI copilot for equity research that reads filings, transcripts, and KPI tables across 100,000+ public companies.

Freemium· Free; Pro $39/mo (annual) or $49/mo; Max and Enterprise API tiers aboveequity-researchearnings-call-analysis
Graphiti preview image
Graphiti logo

Graphiti

RAG · Multi-model
7.2

Open-source temporal knowledge graph framework for building agent memory that updates in real time.

Freemium· Basic: $10 · Pro: $20 · Enterprise: $50agent-memorytemporal-knowledge-graphs
OneKE preview image
OneKE logo

OneKE

RAG · Multi-model (OneKE-13B, LLaMA3, Qwen2.5, GPT, DeepSeek-R1)
7.2

Open-source multi-agent framework for schema-guided knowledge extraction from documents.

Free· Free, MIT-licensed; you pay for LLM API calls or self-hosted computeknowledge-graph-constructionnamed-entity-recognition
Perplexity AI preview image
Perplexity AI logo

Perplexity AI

RAG · Multi-model (Sonar, GPT-4 class, Claude, Gemini)
7.2

Conversational answer engine that cites its sources by default.

Freemium· Basic: $10 · Pro: $30 · Enterprise: Contact salesai-searchresearch
WeKnora preview image
WeKnora logo

WeKnora

RAG · Multi-model
7.2

Tencent's open-source RAG framework that turns raw documents into a queryable knowledge base, ReAct agent, and self-maintaining wiki.

Free· Free, open-source (self-hosted)document-qaenterprise-knowledge-base
BGE (BAAI General Embedding) preview image
BGE (BAAI General Embedding) logo

BGE (BAAI General Embedding)

RAG · BGE / bge-m3 / bge-reranker
7.1

Open-source embedding and reranker models from BAAI that anchor a huge share of production RAG stacks.

Free· Free, open-source (MIT-style license); self-hosted inference cost onlysemantic-searchrag-retrieval
Elicit preview image
Elicit logo

Elicit

RAG · Claude Opus 4.5
7.1

AI research assistant that searches, screens, and extracts data from 138M+ academic papers at scale.

Freemium· Basic: Free · Plus: $11 · Pro: $39 · Scale: $89 · Enterprise: Customliterature-reviewsystematic-review
FutureHouse Platform preview image
FutureHouse Platform logo

FutureHouse Platform

RAG · Multi-model
7.1

Multi-agent AI research stack for scientists, with retrieval over 175M+ papers, patents, and trials.

Freemium· Basic: $10 · Pro: $25 · Enterprise: Contact salesscientific-literature-searchautonomous-research-agent
LangExtract preview image
LangExtract logo

LangExtract

RAG · Multi-model (Gemini, GPT-4/4o, Ollama-hosted local models)
7.1

Google's open-source Python library for LLM-driven structured extraction from unstructured text, with source-grounded outputs.

Free· Library is free (Apache-2.0); LLM API costs depend on chosen backendstructured-extractiondocument-parsing
OpenDataLoader PDF preview image
OpenDataLoader PDF logo

OpenDataLoader PDF

RAG
7.1

Open-source PDF parser built for RAG pipelines, with reading-order detection, table extraction, and bounding-box citations.

Freemium· Free (Apache 2.0); enterprise tier for PDF/UA export and visual editorpdf-parsingrag-preprocessing
PostgresML preview image
PostgresML logo

PostgresML

RAG · Multi-model (Llama, Mistral, open-source embeddings)
7.1

PostgreSQL extension that runs embeddings, vector search, and LLM inference inside your database.

Freemium· Serverless: From $7.50 per query hour · Dedicated: From $0.60 per instance hour · Enterprise: Custom pricingvector-searchrag
Rivestack preview image
Rivestack logo

Rivestack

RAG · OpenAI embeddings (auto-embeddings)
7.1

Managed Postgres with pgvector on dedicated NVMe, pitched as a cheaper RAG backend than Pinecone or Supabase.

Freemium· Free: $0/month · Solo: $29/month · HA Cluster: Starting at $49/node/monthrag-backendvector-search
SiteGPT preview image
SiteGPT logo

SiteGPT

RAG · GPT-4 (per testimonials; not publicly specified)
7.1

Custom GPT-powered support chatbots trained on your website content and docs.

Freemium· Starter: $39 · Growth: $79 · Scale: $259 · Enterprise: Customcustomer-supportwebsite-chatbot
SiteSpeakAI preview image
SiteSpeakAI logo

SiteSpeakAI

RAG · Multi-model
7.1

Custom-trained chatbot that turns your website, docs, and PDFs into a multilingual support and lead-gen agent.

Freemium· Free: Free · Starter: per month · Pro: per month · Growth: per month · Business: per monthwebsite-chatbotcustomer-support
UltraRAG preview image
UltraRAG logo

UltraRAG

RAG · Multi-model (MiniCPM-Embedding-Light, AgentCPM-Report, BYO LLM)
7.1

Low-code, YAML-driven RAG pipeline orchestrator with a visual UI for building and demoing retrieval systems.

Free· Open source; self-hostedrag-pipelinesknowledge-base-qa
alphaXiv preview image
alphaXiv logo

alphaXiv

RAG · Multi-model
7.0

AI reading layer over arXiv with grounded Q&A, auto-summaries, and line-by-line discussion on every preprint.

Free· Free, no signup requiredpaper-qaliterature-review
GaliChat preview image
GaliChat logo

GaliChat

RAG
7.0

No-code AI chatbot builder that trains on your website content for support and lead capture.

Freemium· pro: $19 · business: $99 · enterprise: $999 · Free Demo: Freecustomer-supportlead-generation