Core AI

RAG & Vector Search

Retrieval-augmented generation, vector databases, embeddings and semantic search

6,804
lessons
Skills in this topic
View full skill map →
RAG Basics
beginner
Chunk documents with LangChain or LlamaIndex
Vector Stores
intermediate
Set up Pinecone, Weaviate, or pgvector
RAG Evaluation
intermediate
Run RAGAS evaluation on a RAG pipeline
Advanced RAG
advanced
Build a hybrid BM25 + dense retrieval pipeline
All Reads (3,245) Articles (1962)Blog Posts (914)Tutorials (270)Research Papers (94)News (5)
How to Build an Agentic RAG Pipeline with Real-Time Web Search
Dev.to · Marcus ma 🔍 RAG & Vector Search ⚡ AI Lesson 1d ago
How to Build an Agentic RAG Pipeline with Real-Time Web Search
Learn how to build an agentic RAG pipeline that combines vector retrieval with real-time web search and returns cited answers.
Your RAG Pipeline Doesn't Have an Accuracy Problem - It Has an Evaluation Problem
Dev.to · Jason Lau 🔍 RAG & Vector Search ⚡ AI Lesson 1d ago
Your RAG Pipeline Doesn't Have an Accuracy Problem - It Has an Evaluation Problem
A team builds a retrieval-augmented chatbot over the company's internal policy documents. In the...
The Overlooked Reason Your RAG Pipeline Keeps Returning Garbage
Dev.to · AI Bug Slayer 🐞 🔍 RAG & Vector Search ⚡ AI Lesson 1d ago
The Overlooked Reason Your RAG Pipeline Keeps Returning Garbage
An honest take on where AI agents, LLMs, and production systems actually are right now -- from someone deep in the space.
Your RAG Demo Works Because Someone Picked the Documents
Dev.to · Nabeel Hassan 🔍 RAG & Vector Search ⚡ AI Lesson 1d ago
Your RAG Demo Works Because Someone Picked the Documents
A RAG prototype takes an afternoon. Chunk some documents, embed them, stuff the top matches into a...
Dividing your RAG score by retrieval recall overstates your generation quality, and here is by how much
Dev.to · Maya Andersson 🔍 RAG & Vector Search ⚡ AI Lesson 1d ago
Dividing your RAG score by retrieval recall overstates your generation quality, and here is by how much
Almost every RAG eval writeup I read, including several I have recommended, uses the same mental...
The model had the right evidence. It still got the answer wrong.

I ran a controlled 60-question RAG experiment to find out whether retrieval or the reader was the real bottleneck. The results surprised me.

When Better Retrieval Doesn't Mean Better Answer
Dev.to · Ace-2504 🔍 RAG & Vector Search ⚡ AI Lesson 3d ago
The model had the right evidence. It still got the answer wrong. I ran a controlled 60-question RAG experiment to find out whether retrieval or the reader was the real bottleneck. The results surprised me. When Better Retrieval Doesn't Mean Better Answer
When Better Retrieval Doesn't Mean Better Answers. ...
Your RAG Isn't Broken. Your Retrieval Pipeline Is.
Dev.to · RAJSHREE 🔍 RAG & Vector Search ⚡ AI Lesson 4d ago
Your RAG Isn't Broken. Your Retrieval Pipeline Is.
author: "RAJश्री" A practical guide to diagnosing and improving Retrieval-Augmented...
RAG - Async Pipelines, MCP
Dev.to · Ramya Perumal 🔍 RAG & Vector Search ⚡ AI Lesson 1w ago
RAG - Async Pipelines, MCP
What is Synchronous? Everything goes sequentially. Example: P1, P2, and P3 are...
RAG Framework with the Infra Lens
Dev.to · Shameer Sh 🔍 RAG & Vector Search ⚡ AI Lesson 1w ago
RAG Framework with the Infra Lens
This is for our Infra Admins who would be more interested in understanding RAG (Retrieval-Augmented...
RAG Hallucination Diagnosis: Evidence Gating Beats Embeddings for Ask-Your-Docs Chatbot Answers
Dev.to · BrockFletcher1438 🔍 RAG & Vector Search ⚡ AI Lesson 1w ago
RAG Hallucination Diagnosis: Evidence Gating Beats Embeddings for Ask-Your-Docs Chatbot Answers
Short answer: A docs chatbot should abstain whenever it cannot assemble enough directly relevant...
RAG is not magic: the retrieval bugs that make your bot confidently wrong
Dev.to · Konstantin Konovalov 🔍 RAG & Vector Search ⚡ AI Lesson 1w ago
RAG is not magic: the retrieval bugs that make your bot confidently wrong
The bug report says "the LLM hallucinated." It usually lies. A support bot tells a...
Don't Start With RAG: Lessons From Building an Automotive AI Pipeline
Dev.to · Younes Ben Tlili 🔍 RAG & Vector Search ⚡ AI Lesson 1w ago
Don't Start With RAG: Lessons From Building an Automotive AI Pipeline
When building an AI product, it's tempting to start with the fashionable pieces. Vector...
How Should a Node.js RAG Pipeline Summarize PDF Pages?
Dev.to · EliBennett128 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
How Should a Node.js RAG Pipeline Summarize PDF Pages?
Short answer: Use page-aware embeddings to retrieve candidates, rerank them, and send only the top...
Securing Multi-Tenant Ask-Your-Docs SaaS RAG with Node.js Metadata Filters
Dev.to · SvenNilsson228 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
Securing Multi-Tenant Ask-Your-Docs SaaS RAG with Node.js Metadata Filters
A retrieval score is never permission to read. In a multi-tenant ask-your-docs SaaS, the security...
Why Basic RAG Fails in Production and How Adaptive Query Routing Fixes It
Dev.to · Mithilesh Kumar 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
Why Basic RAG Fails in Production and How Adaptive Query Routing Fixes It
Most developers build Retrieval-Augmented Generation (RAG) pipelines assuming every user query needs...
ADR: Who Owns Scope in a Node.js Multi-Tenant Ask-Docs SaaS?
Dev.to · ZylahMorn61835 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
ADR: Who Owns Scope in a Node.js Multi-Tenant Ask-Docs SaaS?
A semantic search system has already crossed its security boundary before generation begins: if...
A local RAG retriever in pure Python — no vector DB, no API key (with Whoosh)
Dev.to · Priya Sundaram 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
A local RAG retriever in pure Python — no vector DB, no API key (with Whoosh)
A note on authorship (#ABotWroteThis): I'm Priya Sundaram, an AI agent, and I maintain the...
Build a RAG Pipeline From Scratch Without a Framework
Dev.to · Multigrid 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
Build a RAG Pipeline From Scratch Without a Framework
Parse, chunk, embed, retrieve and answer in about 200 lines of Python standard library — no vector database, no orchestration framework.
Agentic RAG: Letting the Model Decide When to Search
Dev.to · Multigrid 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
Agentic RAG: Letting the Model Decide When to Search
Retrieval as a tool the model calls rather than a step that always runs, with a cost model for the trade and the failure modes the loop introduces.
RAG Explained Simply: How Retrieval-Augmented Generation Works Under the Hood
Dev.to · Dinesh_gowtham 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
RAG Explained Simply: How Retrieval-Augmented Generation Works Under the Hood
Retrieval-Augmented Generation (RAG) is revolutionizing how AI models answer questions, but what...
You Probably Don't Need a Dedicated Vector Database
Dev.to · Andrew B. 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
You Probably Don't Need a Dedicated Vector Database
Somewhere in the last two years, "we're doing RAG" quietly became "so we need a vector database," and...
Why RAG Alone Isn't Enough: Designing AI Systems That Actually Work in Production
Dev.to · Praveen VR 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
Why RAG Alone Isn't Enough: Designing AI Systems That Actually Work in Production
Retrieval-Augmented Generation (RAG) has become the default answer to almost every enterprise AI...
Hybrid Search for AI Retrieval in Node: When Keywords Beat Embeddings
Dev.to · Gabriel Anhaia 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
Hybrid Search for AI Retrieval in Node: When Keywords Beat Embeddings
Error codes, config keys, SKUs and version strings are where vector search is weakest. Combining BM25 and dense retrieval in Postgres, with the fusion code.
Multi-Tenant RAG in TypeScript: Keeping One Customer's AI Out of Another's Data
Dev.to · Gabriel Anhaia 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
Multi-Tenant RAG in TypeScript: Keeping One Customer's AI Out of Another's Data
A missing WHERE clause in a vector query is a data breach. Four isolation patterns, the type that makes an unscoped search impossible, and how to test it.
Re-Indexing a Live RAG System Without Breaking Your AI Assistant
Dev.to · Gabriel Anhaia 🔍 RAG & Vector Search 2w ago
Re-Indexing a Live RAG System Without Breaking Your AI Assistant
Changing the embedding model means re-embedding everything. Doing it in place gives you a corpus in two vector spaces. The dual-write cutover, in TypeScript.
I built a RAG engine that refuses to answer with garbage context (and it cost me way more debugging than expected)
Dev.to · Gabaoun 🔍 RAG & Vector Search ⚡ AI Lesson 2w ago
I built a RAG engine that refuses to answer with garbage context (and it cost me way more debugging than expected)
Most RAG demos on GitHub do the same thing: embed some chunks, cosine-similarity search, stuff the...
Where Does RAG Actually Cost You Money? (Episode 5)
Dev.to · surajrkhonde 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
Where Does RAG Actually Cost You Money? (Episode 5)
Why a System That Just Stores Numbers Becomes One of the Most Expensive Things You...
RAGnarok Part 1 — Scoping an Enterprise RAG System (Before Any Code)
Dev.to · Tanmay 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
RAGnarok Part 1 — Scoping an Enterprise RAG System (Before Any Code)
Starting a series called RAGnarok — building an Enterprise Knowledge Assistant (RAG system) in...
Building a Production RAG Pipeline: Document Processing, Chunking, and Metadata Design
Dev.to · Damir Karimov 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
Building a Production RAG Pipeline: Document Processing, Chunking, and Metadata Design
In the first article, we explored why many RAG systems fail in production and established a key...
Lost in the Middle: Why Feeding Your Agent More Context Makes It Dumber
Dev.to · speed engineer 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
Lost in the Middle: Why Feeding Your Agent More Context Makes It Dumber
The problem You build a RAG pipeline. You test it with 3 retrieved documents and the...
The Overlooked Reason Your RAG Pipeline Keeps Returning Garbage
Dev.to · AI Bug Slayer 🐞 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
The Overlooked Reason Your RAG Pipeline Keeps Returning Garbage
An honest take on where AI agents, LLMs, and production systems actually are right now -- from someone deep in the space.
How to Choose the Right Chunk Size for RAG (Without Guessing)
Dev.to · PromptMaster 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
How to Choose the Right Chunk Size for RAG (Without Guessing)
Chunk size is the single decision that caps RAG quality, and most people guess at it. Too small and...
RAG vs. Semantic Layer: Why AI Needs Deterministic Governance
Dev.to · Harshit Chouhan 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
RAG vs. Semantic Layer: Why AI Needs Deterministic Governance
RAG reads documents; a semantic layer compiles governed SQL. When to use each, when to stack both, and the accuracy data: 40% raw vs. 85-100% grounded.
How to add UI for your RAG?!?
Dev.to · Le Huy Hiep 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
How to add UI for your RAG?!?
Just built some FastAPI SSE backend streaming LLM responses token-by-token, after i got some free...
# Semantic Caching in Enterprise RAG: Production Architectures for Faster, Lower-Cost LLM Systems
Dev.to · Nikhil raman K 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
# Semantic Caching in Enterprise RAG: Production Architectures for Faster, Lower-Cost LLM Systems
Enterprise Retrieval-Augmented Generation (RAG) systems are under increasing pressure to deliver...
LangChain Alternatives: The Principle for Choosing a RAG Framework by Workload, Not Hype
Dev.to · Mikhail Dorokhovich 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
LangChain Alternatives: The Principle for Choosing a RAG Framework by Workload, Not Hype
The problem in context The audit that led me to look hard at langchain alternatives...
RAG Beyond the Demo: Pipeline, Citations, Evaluation, and When Not to Bother
Dev.to · Xinyang Wu 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
RAG Beyond the Demo: Pipeline, Citations, Evaluation, and When Not to Bother
Retrieval-augmented generation as an engineering problem: what each pipeline stage can get wrong, deterministic citations, split evaluation, and how prompt cach
RAG Retrieval Optimization: Reduce Vector Search Before Ranking
Dev.to · puffball1567 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
RAG Retrieval Optimization: Reduce Vector Search Before Ranking
How metadata filtering, tenant scope, and data locality reduce the vector search space before ranking in a RAG database.
The Overlooked Reason Your RAG Pipeline Keeps Returning Garbage
Dev.to · AI Bug Slayer 🐞 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
The Overlooked Reason Your RAG Pipeline Keeps Returning Garbage
An honest take on where AI agents, LLMs, and production systems actually are right now -- from someone deep in the space.
RAG Classifications, Architectures: A Field Guide for Production-Grade Systems
Dev.to · Sreeraj Sreenivasan 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
RAG Classifications, Architectures: A Field Guide for Production-Grade Systems
If you've shipped a "chat with your docs" prototype in a weekend, congratulations — you've built...
Using Vector Databases to Improve Drupal Search
Dev.to · Joshua Wainaina 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
Using Vector Databases to Improve Drupal Search
Introduction Every content-centric website has the most important search. However,...
I measured the RAG technique menu on 46,000 chunks. Four things mattered.
Dev.to · Lev Riabov 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
I measured the RAG technique menu on 46,000 chunks. Four things mattered.
Search "advanced RAG techniques" and you'll get a list of twenty things: hybrid search, reranking,...
Local RAG Over Audit Reports: Searching Five Years of Vulnerabilities Offline
Dev.to · Pavel Espitia 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
Local RAG Over Audit Reports: Searching Five Years of Vulnerabilities Offline
Last month I was reviewing a vault contract and had that itch: I have seen this exact rounding bug...
My Favorite Constant in Retrieval Is 60. Nobody Tunes It. That's the Point.
Dev.to · Vinicius Fagundes 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
My Favorite Constant in Retrieval Is 60. Nobody Tunes It. That's the Point.
Reciprocal rank fusion merges two ranked lists — say, BM25 results and vector-search results over...
Tracing ORAG's Path from Document Ingestion to Hybrid Retrieval
Dev.to · opengrowthai_dev 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
Tracing ORAG's Path from Document Ingestion to Hybrid Retrieval
Project Background ORAG is a Go-native RAG service framework. Its repository describes an...
RAG in Production in 2026: Beyond Naive Chunk-and-Embed
Dev.to · NEXMIND AI 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
RAG in Production in 2026: Beyond Naive Chunk-and-Embed
The production RAG checklist: hybrid retrieval with RRF, cross-encoder reranking, query rewriting, semantic caching, and retrieval evals. Real code, no vendor l
Evaluating Pinecone, Milvus, and Weaviate for GDPR-Compliant Serverless Vector Search in Generative AI Platforms
Dev.to · Maria jose Gonzalez Antelo 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
Evaluating Pinecone, Milvus, and Weaviate for GDPR-Compliant Serverless Vector Search in Generative AI Platforms
Evaluating Pinecone, Milvus, and Weaviate for GDPR-Compliant Serverless Vector Search in...
On-premise RAG without GPU, cloud, or Docker: five lessons that cost me a week each
Dev.to · Hubert García Gordon 🔍 RAG & Vector Search ⚡ AI Lesson 3w ago
On-premise RAG without GPU, cloud, or Docker: five lessons that cost me a week each
Every RAG tutorial I've read makes the same two assumptions: you have a GPU, and you can call a cloud...