Skip to main content
📖 The AI Tool Bible

RAG

Retrieval-augmented generation, vector stores, indexers.

95 tools

Why it matters

RAG isn't a model, it's an architecture — retrieve, augment, generate. The choice is between frameworks that orchestrate the retrieval and the vector stores underneath.

What's in here

Includes RAG frameworks (LlamaIndex, LangChain), managed vector databases (Pinecone), open-source vector stores (Weaviate, Chroma, Vespa), and hybrid-search engines.

How to pick

Pick LlamaIndex when retrieval quality is the bottleneck. Pick Pinecone for zero-ops production. Pick Weaviate or Chroma for self-hosted or budget-conscious. Pick Vespa at scale beyond a few million docs.

PI

Pinecone

Featured
RAG · Hosted vector DB (not an LLM)
8.8

Managed vector database for production-scale similarity search.

Freemium· Starter: Free · Builder: $20/month flat · Standard: $50/month min. usage · Enterprise: $500/month min. usagemanaged vector DBproduction RAG
LL

LlamaIndex

Featured
RAG · BYO (Claude / GPT / open)
8.7

Data framework for connecting LLMs to your data.

Freemium· Free open-source; LlamaCloud paidRAGdata ingestion
EV

Elasticsearch Vector Search

RAG · BYO embeddings (OpenAI, Cohere, Hugging Face, Mistral, Bedrock, Vertex, Azure) plus Elastic's built-in ELSER sparse model and E5 dense model
8.7

Hybrid vector + keyword search in the enterprise-grade Elasticsearch engine

Freemium· Resource based pricing: Pay as you go (monthly) or prepaid · Usage based pricing: Pay as you go (monthly) or prepaid · License based pricing: ?RAG chatbot over enterprise docsHybrid semantic + keyword product search
SC

Snowflake Cortex

RAG · Anthropic Claude, Meta Llama, Mistral Large 2, Snowflake Arctic
8.7

Generative AI and RAG built into the Snowflake data cloud

Enterprise· Standard: Contact sales · Enterprise: Contact sales · Business Critical: Contact sales · Virtual Private Snowflake: Contact salesEnterprise RAG chatbot over governed dataNatural-language SQL for business analysts
DA

DataStax Astra DB

RAG · Bring-your-own embeddings; integrates with OpenAI, Cohere, Hugging Face, Mistral, NVIDIA NIM, and Vertex AI via server-side vectorize
8.6

Serverless vector and document database for production RAG and AI agents

Freemium· Small On-Demand: Contact sales · Medium (Balanced): Contact sales · Medium (Storage Optimized): Contact sales · Large (Balanced): Contact sales · Large (Storage Optimized): Contact salesRAG chatbot over enterprise documentsAgent long-term memory store
MA

MongoDB Atlas Vector Search

RAG · Bring-your-own embeddings (OpenAI, Cohere, open models); native Voyage AI embeddings and rerankers
8.6

Vector search built into the operational database you're already using.

Freemium· Free: $0 · Flex: Up to $30 · Dedicated: Starts at $56.94RAG over enterprise documentsProduct and content recommendation engines
QU

Quivr

RAG · Multi-model (OpenAI, Anthropic, Mistral, Gemma)
8.4

Open-source RAG framework for building custom AI assistants over your own documents in a few lines of Python.

Free· Open source (pip install quivr-core); pay only for LLM/vector-store usagedocument-qacustom-knowledge-base
WE

Weaviate

RAG · Hosted vector DB (not an LLM)
8.4

Open-source vector DB with hybrid search and modules.

Freemium· Free: $0 · Flex: $45 · Premium: $400self-hosted RAGhybrid search
LA

LangChain

RAG · BYO (any major LLM)
8.3

The broad LLM application framework — chains, agents, retrievers.

Freemium· Free open-source; LangSmith paidgeneral LLM appsRAG
VA

Vanna.ai

RAG · Multi-model (Anthropic, OpenAI, Gemini, Ollama)
8.3

Open-source text-to-SQL agent that learns your schema and writes queries against your real warehouse.

Freemium· Explorer: $50 · Team: $500 · Enterprise: Contact salestext-to-sqlnatural-language-bi
FE

Feast

RAG
8.2

Open-source feature store that serves consistent features to ML training and online inference, with RAG vector search built in.

Free· Free, open source (Apache 2.0); self-hostedfeature-storerag-retrieval
LA

LanceDB

RAG
8.2

Open-source multimodal lakehouse and vector database built for AI training and retrieval at petabyte scale.

Freemium· Open-source free; LanceDB Cloud and Enterprise via contact salesvector-searchrag
SC

Scite

RAG · Multi-model
8.2

AI research assistant that grades citations as supporting, contrasting, or mentioning across 1.6B citation statements.

Freemium· Basic: $20 · Pro: $50 · Team: $50 · Enterprise: Contact usliterature-reviewcitation-analysis
VE

Vespa

RAG · Hosted search engine (not an LLM)
8.2

Yahoo's open-source search engine with vector + sparse retrieval.

Freemium· Free open-source; Vespa Cloud paidlarge-scale searchranking
CH

Chroma

RAG · Hosted vector DB (not an LLM)
8.1

Embedded, developer-friendly vector store for Python.

Freemium· Starter: $0 · Team: $250 · Enterprise: Customprototypingembedded RAG
CU

Cube

RAG · Multi-model
8.1

Semantic layer that grounds LLM agents in your real business metrics instead of letting them hallucinate SQL.

Freemium· Cube Core open source; Cube Cloud paid, contact salessemantic-layerembedded-analytics
DV

Databricks Vector Search

RAG · Multi-model (BYO embeddings or Databricks-hosted)
8.1

Managed hybrid vector search that lives inside the Databricks lakehouse and auto-syncs with your source tables.

Enterprise· Standard: $605 · Storage Optimized: $922rag-retrievalhybrid-search
NO

NotebookLM

RAG · Gemini 2.5
8.1

Google's source-grounded research notebook that turns your documents into chats, briefs, and AI-hosted podcasts.

Freemium· Free tier; Plus via Google One AI Premium ($19.99/mo) or Workspace add-ondocument Q&Aresearch synthesis
RA

RAGFlow

RAG · Multi-model
8.1

Open-source RAG engine with deep document parsing, hybrid search, and visual agent orchestration.

Freemium· Free tier; Starter $29/mo; Pro $129/mo; Enterprise customdocument-qaenterprise-search
EX

Exa

RAG · Proprietary neural + keyword search
8.0

Web search API built for AI agents, with structured outputs and token-efficient highlights.

Freemium· Free Tier: Free · Search: $7/1k requests · Agent: $0.012–$1.00/run · Contents: $1/1k pages per content type · Deep Search: $12–15/1k requestsagent-web-searchrag-retrieval
FI

Firecrawl

RAG · Claude, Cursor, Windsurf, OpenAI, Gemini
8.0

Web scraping and crawling API that returns LLM-ready markdown, JSON, or structured data from any URL.

Freemium· Free Plan: $0 · Hobby: $16 · Standard: $83 · Growth: $333 · Scale: $599/monthlyweb-scrapingrag-ingestion
WA

Wren AI

RAG · Multi-model (OpenAI, Anthropic, Gemini, self-hosted)
8.0

Open-source GenBI semantic layer that lets AI agents query your warehouse in natural language with governed, accurate SQL.

Freemium· Free: $0 · Essential Cloud: $179 · Enterprise Cloud: $559 · Enterprise Plus: Contact Ustext-to-sqlsemantic-layer
AN

AnythingLLM

RAG · Multi-model
7.9

Open-source desktop and self-hosted app that turns your documents into a private chat-and-agent workspace.

Freemium· Basic: $50/monthly · Pro: $99/monthly · Enterprise: Contact Usdocument-chatprivate-rag
HA

Humata.ai

RAG · Multi-model
7.8

Chat-with-your-documents RAG tool with citation-backed answers across uploaded PDFs and files.

Freemium· Free: $0 · Expert: $9.99 · Team: $49 / user · Enterprise: customdocument-qaresearch-summarization