DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
5 Reasons Your RAG System Will Fail in Production (And the Patterns I Use to Fix Each One)

5 Reasons Your RAG System Will Fail in Production (And the Patterns I Use to Fix Each One)

4
Comments 4
5 min read
CARE Loop: A Human-Centered Framework for Local LLM Development

CARE Loop: A Human-Centered Framework for Local LLM Development

2
Comments
5 min read
Reports of RAG's death have been greatly exaggerated

Reports of RAG's death have been greatly exaggerated

Comments
5 min read
Building Production-Ready RAG: A Complete Architecture Guide

Building Production-Ready RAG: A Complete Architecture Guide

1
Comments
7 min read
RAG Series (19): Incremental Updates — Keeping the Knowledge Base Fresh

RAG Series (19): Incremental Updates — Keeping the Knowledge Base Fresh

2
Comments
6 min read
Llama.cpp's New MTP on MacOS

Llama.cpp's New MTP on MacOS

1
Comments
4 min read
The Amnesia Epidemic: Why the Next Era of Enterprise AI Requires "Hindsight"

The Amnesia Epidemic: Why the Next Era of Enterprise AI Requires "Hindsight"

Comments
4 min read
Build an End-to-End Smart Semantic Search App Using LangChain

Build an End-to-End Smart Semantic Search App Using LangChain

Comments
4 min read
Building a Biomedical GraphRAG Inference System: Comparing LLM-Only, Basic RAG, and GraphRAG Pipelines

Building a Biomedical GraphRAG Inference System: Comparing LLM-Only, Basic RAG, and GraphRAG Pipelines

1
Comments
3 min read
How We Built CyberGraph RAG: A 3.5M Token Cybersecurity GraphRAG System with TigerGraph

How We Built CyberGraph RAG: A 3.5M Token Cybersecurity GraphRAG System with TigerGraph

2
Comments
3 min read
Applying RAG Architectures to Travel Knowledge Bases: A Practitioner's View

Applying RAG Architectures to Travel Knowledge Bases: A Practitioner's View

Comments
7 min read
Tackle High Token Usage with GraphRAG

Tackle High Token Usage with GraphRAG

2
Comments
4 min read
LLM Prompting, AI-Generated Code Discussions & Python Workflow Automation

LLM Prompting, AI-Generated Code Discussions & Python Workflow Automation

Comments
3 min read
Building KernelMind, A Code-Aware Github Companion

Building KernelMind, A Code-Aware Github Companion

1
Comments 2
6 min read
RAG- Understanding of Embedding

RAG- Understanding of Embedding

Comments
3 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.