DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Rebuilding the Cerebras Knowledge Base: planner, tools, and synthesis

Rebuilding the Cerebras Knowledge Base: planner, tools, and synthesis

Comments
6 min read
Rebuilding the Cerebras Knowledge Base: an LLM reranker

Rebuilding the Cerebras Knowledge Base: an LLM reranker

Comments
5 min read
LLMOps for Production RAG: Observability & Cost Guide

LLMOps for Production RAG: Observability & Cost Guide

2
Comments 1
4 min read
Eval-First RAG: Use Separate Scores to Triage Failures

Eval-First RAG: Use Separate Scores to Triage Failures

Comments 1
5 min read
Building a RAG-Powered Code Review Assistant with PHP, Ollama, and Qdrant

Building a RAG-Powered Code Review Assistant with PHP, Ollama, and Qdrant

Comments
21 min read
Construire un assistant de code review RAG avec PHP, Ollama et Qdrant

Construire un assistant de code review RAG avec PHP, Ollama et Qdrant

Comments
28 min read
Qdrant TypeScript in 2026: use query(), not stale search() examples

Qdrant TypeScript in 2026: use query(), not stale search() examples

Comments 1
3 min read
How to Tell When Your LLM's Knowledge Actually Stops

How to Tell When Your LLM's Knowledge Actually Stops

Comments
2 min read
RAG vs. Fine-Tuning: The AI Engineer's Decision Framework

RAG vs. Fine-Tuning: The AI Engineer's Decision Framework

4
Comments 1
3 min read
RAGBench: Stop Guessing, Start Measuring Your RAG Pipeline

RAGBench: Stop Guessing, Start Measuring Your RAG Pipeline

Comments 2
1 min read
Architecture of an Auditable Al Chatbot: Multi-Agent Routing with Open Knowledge Format

Architecture of an Auditable Al Chatbot: Multi-Agent Routing with Open Knowledge Format

Comments 2
9 min read
How MCP fetch servers keep getting SSRF wrong (and how I tried not to)

How MCP fetch servers keep getting SSRF wrong (and how I tried not to)

1
Comments
6 min read
Securing Enterprise AI Pipelines: Using Azure Key Vault and AES Encryption for Sensitive Documents in Snowflake/Databricks AI Workloads

Securing Enterprise AI Pipelines: Using Azure Key Vault and AES Encryption for Sensitive Documents in Snowflake/Databricks AI Workloads

Comments 1
7 min read
BM25 Length Normalization: Why Long RAG Chunks Never Rank

BM25 Length Normalization: Why Long RAG Chunks Never Rank

Comments
7 min read
Firecrawl vs. Jina Reader vs. Tavily: Picking the Right Web-to-LLM Tool (and When You Need None of Them)

Firecrawl vs. Jina Reader vs. Tavily: Picking the Right Web-to-LLM Tool (and When You Need None of Them)

1
Comments
5 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.