DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
How We Automated Hallucination Detection in Enterprise RAG Pipelines

How We Automated Hallucination Detection in Enterprise RAG Pipelines

Comments
1 min read
I tested 7 vector databases for my RAG stack in 2026, here's the one nobody is talking about (yet)

I tested 7 vector databases for my RAG stack in 2026, here's the one nobody is talking about (yet)

15
Comments 2
2 min read
I Built an AI Chatbot Into My Portfolio Website Using AWS Bedrock — Here's Exactly How

I Built an AI Chatbot Into My Portfolio Website Using AWS Bedrock — Here's Exactly How

1
Comments
10 min read
RAG vs MCP is the wrong debate — here's the right framing for production AI systems

RAG vs MCP is the wrong debate — here's the right framing for production AI systems

Comments
4 min read
When NOT to use RAG (lessons from building a Claude-powered support bot)

When NOT to use RAG (lessons from building a Claude-powered support bot)

Comments
4 min read
Optimizing LLM Workflows: Claude for Evaluation, Blender Integration & Token Efficiency

Optimizing LLM Workflows: Claude for Evaluation, Blender Integration & Token Efficiency

Comments
3 min read
How a Single PDF Can Poison 100 RAG Systems: The Vulnerability We Aren't Talking About

How a Single PDF Can Poison 100 RAG Systems: The Vulnerability We Aren't Talking About

Comments
4 min read
40 Days Training on RAG

40 Days Training on RAG

Comments
2 min read
🎓 Session 1: Hello World of RAG + Introduction & Need of RAG

🎓 Session 1: Hello World of RAG + Introduction & Need of RAG

Comments
5 min read
I Built NativeLM for Android (And Bypassed OEM RAM Lies to Do It)

I Built NativeLM for Android (And Bypassed OEM RAM Lies to Do It)

1
Comments
4 min read
In This Memory Test, Relevance Wasn't Authority

In This Memory Test, Relevance Wasn't Authority

1
Comments 6
9 min read
RAG on call transcripts: utterance-aware chunking and hybrid retrieval in production

RAG on call transcripts: utterance-aware chunking and hybrid retrieval in production

Comments
6 min read
How to detect AI hallucinations inside n8n — RagMetrics node walkthrough

How to detect AI hallucinations inside n8n — RagMetrics node walkthrough

Comments
2 min read
How I Stopped Worrying and Made RAG Work in User-Message (80% Compliance, Solo)

How I Stopped Worrying and Made RAG Work in User-Message (80% Compliance, Solo)

Comments
1 min read
How to Fine-Tune Llama 3.1 8B for Under $5 Using QLoRA in 2026 – A Practical Guide

How to Fine-Tune Llama 3.1 8B for Under $5 Using QLoRA in 2026 – A Practical Guide

1
Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.