DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Context Arithmetic: A 5-Stage Retrieval Pipeline for Voice Agents (500K docs to 400 tokens in <200ms)

Context Arithmetic: A 5-Stage Retrieval Pipeline for Voice Agents (500K docs to 400 tokens in <200ms)

2
Comments
5 min read
Your RAG Filter Runs Too Late. Build a Tenant-Safe Retriever in TypeScript.

Your RAG Filter Runs Too Late. Build a Tenant-Safe Retriever in TypeScript.

Comments
6 min read
How to Build Financial Dashboard Pagination with Traceable Retrieval Contracts

How to Build Financial Dashboard Pagination with Traceable Retrieval Contracts

Comments
6 min read
My RAG Eval Passed. The Citations Were Still Wrong

My RAG Eval Passed. The Citations Were Still Wrong

Comments
3 min read
Retrieval confidence can't tell your RAG chatbot when the answer is missing

Retrieval confidence can't tell your RAG chatbot when the answer is missing

2
Comments 2
7 min read
Building a RAG App That Knows When to Say "I Don't Know"

Building a RAG App That Knows When to Say "I Don't Know"

Comments
2 min read
EduNode: Building a Multilingual Adaptive AI Tutor with Gemma and RAG

EduNode: Building a Multilingual Adaptive AI Tutor with Gemma and RAG

1
Comments
2 min read
Close the Loop: Chatbot Analytics That Turn Misses Into Fixes

Close the Loop: Chatbot Analytics That Turn Misses Into Fixes

Comments
3 min read
OpsPilot AI: Building RAG and Agents Without Giving the LLM Authority

OpsPilot AI: Building RAG and Agents Without Giving the LLM Authority

Comments
11 min read
UNREAL Unifies Retrieval & Long‑Context—Cut Latency 50% in One Pass!

UNREAL Unifies Retrieval & Long‑Context—Cut Latency 50% in One Pass!

Comments
7 min read
How AI search picks fragments: query fan-out and RAG explained

How AI search picks fragments: query fan-out and RAG explained

Comments
11 min read
EmbeddingGemma 2: Building Local Multimodal RAG with Python

EmbeddingGemma 2: Building Local Multimodal RAG with Python

Comments
5 min read
My Eval Said RAG Made Things Up. My Eval Was Wrong.

My Eval Said RAG Made Things Up. My Eval Was Wrong.

3
Comments
6 min read
Travel Listing Retrieval Architecture: Latency Budgets for Grounded Itinerary Answers

Travel Listing Retrieval Architecture: Latency Budgets for Grounded Itinerary Answers

Comments
7 min read
Voyage AI vs Cohere Embed v4 vs Nemotron 3 Embed: Production RAG Trade-offs

Voyage AI vs Cohere Embed v4 vs Nemotron 3 Embed: Production RAG Trade-offs

Comments
9 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.