DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Revisiting My Phone AI After Gemma 4: The Upgrade I Didn't Know I Needed

Revisiting My Phone AI After Gemma 4: The Upgrade I Didn't Know I Needed

Comments
4 min read
NyayAI: AI-Powered Legal Intelligence for India

NyayAI: AI-Powered Legal Intelligence for India

1
Comments
13 min read
How AI Memory Actually Works: Context Windows and RAG

How AI Memory Actually Works: Context Windows and RAG

Comments
8 min read
Semantic Chunking with Overlap and Section-Awareness: The RAG Tutorial Nobody Wrote

Semantic Chunking with Overlap and Section-Awareness: The RAG Tutorial Nobody Wrote

Comments
8 min read
Stop Benchmarking Embedding Models. 90% of Your Search Quality Lives Upstream.

Stop Benchmarking Embedding Models. 90% of Your Search Quality Lives Upstream.

Comments
4 min read
Applied Claude: Data Recovery, Agent Orchestration, Real-time Content

Applied Claude: Data Recovery, Agent Orchestration, Real-time Content

Comments
3 min read
Chunking Strategies for LLM Applications: A Practical Guide to Better RAG Systems

Chunking Strategies for LLM Applications: A Practical Guide to Better RAG Systems

1
Comments
4 min read
Hybrid Search in RAG: Why Neither Keyword Search Nor Semantic Search Alone Is Good Enough

Hybrid Search in RAG: Why Neither Keyword Search Nor Semantic Search Alone Is Good Enough

Comments
13 min read
Architecting for Speed and Precision: My Blueprint for a Production-Ready RAG System

Architecting for Speed and Precision: My Blueprint for a Production-Ready RAG System

1
Comments
4 min read
7 Production RAG Mistakes I Made (And How to Fix Them)

7 Production RAG Mistakes I Made (And How to Fix Them)

1
Comments
5 min read
Why RAG Pipelines Silently Hallucinate — And The Decay Score That Catches It Before The LLM Does

Why RAG Pipelines Silently Hallucinate — And The Decay Score That Catches It Before The LLM Does

Comments
2 min read
Why Do We Need GraphRAG? — The Evolution from "Search" to "Understanding"

Why Do We Need GraphRAG? — The Evolution from "Search" to "Understanding"

Comments
5 min read
How to Build a RAG Chatbot with Python

How to Build a RAG Chatbot with Python

1
Comments
3 min read
Chunk Overlap: The RAG Parameter Most Teams Pick Wrong

Chunk Overlap: The RAG Parameter Most Teams Pick Wrong

Comments
7 min read
Reranker Selection: Cross-Encoder vs LLM-as-Reranker vs ColBERT: Which Earns Its Latency

Reranker Selection: Cross-Encoder vs LLM-as-Reranker vs ColBERT: Which Earns Its Latency

Comments
9 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.