Skip to main content
📖 The AI Tool Bible

The RAG stack that actually scales

Twelve pieces — embedding, vector store, retrieval, eval, framework — that hold up past a prototype.

The RAG hello-world looks easy. The production RAG stack is a different beast: you need an embedding model, a vector store, a retrieval framework, a reranker, and an eval harness that catches regressions before users do. These twelve tools cover every layer. Pick one from each — don't pick two from the same.

13 tools in this collection
PI

Pinecone

Featured
RAG · Hosted vector DB (not an LLM)
8.8

Managed vector database for production-scale similarity search.

Freemium· Starter: Free · Builder: $20/month flat · Standard: $50/month min. usage · Enterprise: $500/month min. usagemanaged vector DBproduction RAG
WE

Weaviate

RAG · Hosted vector DB (not an LLM)
8.4

Open-source vector DB with hybrid search and modules.

Freemium· Free: $0 · Flex: $45 · Premium: $400self-hosted RAGhybrid search
CH

Chroma

RAG · Hosted vector DB (not an LLM)
8.1

Embedded, developer-friendly vector store for Python.

Freemium· Starter: $0 · Team: $250 · Enterprise: Customprototypingembedded RAG
VE

Vespa

RAG · Hosted search engine (not an LLM)
8.2

Yahoo's open-source search engine with vector + sparse retrieval.

Freemium· Free open-source; Vespa Cloud paidlarge-scale searchranking
LL

LlamaIndex

Featured
RAG · BYO (Claude / GPT / open)
8.7

Data framework for connecting LLMs to your data.

Freemium· Free open-source; LlamaCloud paidRAGdata ingestion
LA

LangChain

RAG · BYO (any major LLM)
8.3

The broad LLM application framework — chains, agents, retrievers.

Freemium· Free open-source; LangSmith paidgeneral LLM appsRAG
FE

Feast

RAG
8.2

Open-source feature store that serves consistent features to ML training and online inference, with RAG vector search built in.

Free· Free, open source (Apache 2.0); self-hostedfeature-storerag-retrieval
RA

RAGFlow

RAG · Multi-model
8.1

Open-source RAG engine with deep document parsing, hybrid search, and visual agent orchestration.

Freemium· Free tier; Starter $29/mo; Pro $129/mo; Enterprise customdocument-qaenterprise-search
HA

Humata.ai

RAG · Multi-model
7.8

Chat-with-your-documents RAG tool with citation-backed answers across uploaded PDFs and files.

Freemium· Free: $0 · Expert: $9.99 · Team: $49 / user · Enterprise: customdocument-qaresearch-summarization
EX

Exa

RAG · Proprietary neural + keyword search
8.0

Web search API built for AI agents, with structured outputs and token-efficient highlights.

Freemium· Free Tier: Free · Search: $7/1k requests · Agent: $0.012–$1.00/run · Contents: $1/1k pages per content type · Deep Search: $12–15/1k requestsagent-web-searchrag-retrieval
NO

NotebookLM

RAG · Gemini 2.5
8.1

Google's source-grounded research notebook that turns your documents into chats, briefs, and AI-hosted podcasts.

Freemium· Free tier; Plus via Google One AI Premium ($19.99/mo) or Workspace add-ondocument Q&Aresearch synthesis
SC

Scite

RAG · Multi-model
8.2

AI research assistant that grades citations as supporting, contrasting, or mentioning across 1.6B citation statements.

Freemium· Basic: $20 · Pro: $50 · Team: $50 · Enterprise: Contact usliterature-reviewcitation-analysis
CO

Cohere

RAG · Command, Embed, Rerank, Transcribe (proprietary)
6.9

Enterprise-grade LLM platform built for private, secure, and customizable deployment.

Enterprise· Embed 4 Small: $2,500 · Embed 4 Medium: $3,250 · Rerank 3.5 Medium: $3,250 · Rerank 4 Fast Medium: $3,250 · Rerank 4 Pro Medium: $3,250enterprise-ragsemantic-search