DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
A Cognitive Benchmark for Code-RAG Retrieval: Part 2 — Why Model Rankings Depend on the Pipeline

A Cognitive Benchmark for Code-RAG Retrieval: Part 2 — Why Model Rankings Depend on the Pipeline

Comments
15 min read
The most underrated feature in a website AI assistant: saying "I don't know"

The most underrated feature in a website AI assistant: saying "I don't know"

Comments
2 min read
Documents Aren't Bags of Chunks

Documents Aren't Bags of Chunks

15
Comments 35
3 min read
Portfolio de IA open source: 20+ proyectos con LLMs self-hosted en producción

Portfolio de IA open source: 20+ proyectos con LLMs self-hosted en producción

Comments
3 min read
Book Library: A Local RAG That Answers From My Own PDFs

Book Library: A Local RAG That Answers From My Own PDFs

Comments
5 min read
Your AI Agent's Memory Has No Expiry Date: I Scored Freshness on a Real Corpus

Your AI Agent's Memory Has No Expiry Date: I Scored Freshness on a Real Corpus

Comments
14 min read
Building a Four-Tier Parallel RAG Pipeline with Gemini

Building a Four-Tier Parallel RAG Pipeline with Gemini

2
Comments 4
2 min read
RAG - Semantic Caching

RAG - Semantic Caching

Comments
3 min read
Optimizing RAG Pipelines, Migrating AI Agents, and LLM-Powered Troubleshooting

Optimizing RAG Pipelines, Migrating AI Agents, and LLM-Powered Troubleshooting

Comments
3 min read
Building a RAG System from Scratch — AI Agents: Memory, Planning, and Multi-Step Reasoning

Building a RAG System from Scratch — AI Agents: Memory, Planning, and Multi-Step Reasoning

Comments 1
7 min read
Your RAG System Is Lying To You About That Table

Your RAG System Is Lying To You About That Table

13
Comments 2
4 min read
A Chinese 8B model beat the Western 8B models at Japanese RAG. I still wouldn't put it in the default deployment — and that distinction is the point.

A Chinese 8B model beat the Western 8B models at Japanese RAG. I still wouldn't put it in the default deployment — and that distinction is the point.

Comments
4 min read
Part 5 — Installing a Black Box Recorder in Your RAG System: 4-Layer Metadata + 3-Level Verification, Root Cause in 5 Minutes

Part 5 — Installing a Black Box Recorder in Your RAG System: 4-Layer Metadata + 3-Level Verification, Root Cause in 5 Minutes

6
Comments
9 min read
Part 2 — Why Does One System Need Three Chunking Strategies? And One Document Type Shouldn't Be Chunked At All

Part 2 — Why Does One System Need Three Chunking Strategies? And One Document Type Shouldn't Be Chunked At All

6
Comments
9 min read
Part 3 — Vector Retrieval in Domain-Specific Terminology Scenarios: From Model Selection to Dual Validation

Part 3 — Vector Retrieval in Domain-Specific Terminology Scenarios: From Model Selection to Dual Validation

6
Comments
10 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.