DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Build a Semantic Cache for Your LLM App in 40 Lines of Python (And Cut Costs by Half)

Build a Semantic Cache for Your LLM App in 40 Lines of Python (And Cut Costs by Half)

1
Comments 3
5 min read
RAG Beyond the Demo: Pipeline, Citations, Evaluation, and When Not to Bother

RAG Beyond the Demo: Pipeline, Citations, Evaluation, and When Not to Bother

Comments 2
7 min read
RAG Retrieval Optimization: Reduce Vector Search Before Ranking

RAG Retrieval Optimization: Reduce Vector Search Before Ranking

1
Comments 2
5 min read
LangChain Alternatives: The Principle for Choosing a RAG Framework by Workload, Not Hype

LangChain Alternatives: The Principle for Choosing a RAG Framework by Workload, Not Hype

Comments 2
6 min read
Why Kimi K3 Still Can't Do What Einstein Did

RAG surfaces echoes, but misses paradigm shifts

Why Kimi K3 Still Can't Do What Einstein Did

35
Comments 22
4 min read
Five Bugs Deep in an AI Memory Layer: My Week with Cognee

Five Bugs Deep in an AI Memory Layer: My Week with Cognee

1
Comments
4 min read
Deploy AI agents in 5 lines of code.

Deploy AI agents in 5 lines of code.

Comments
3 min read
7 Cheapest Web Search APIs for AI Agents in 2026, Ranked

7 Cheapest Web Search APIs for AI Agents in 2026, Ranked

Comments
4 min read
On-premise RAG without GPU, cloud, or Docker: five lessons that cost me a week each

On-premise RAG without GPU, cloud, or Docker: five lessons that cost me a week each

2
Comments 15
8 min read
# We Keep Teaching AI to Retrieve Information. What If We Taught It to Understand It Instead?

# We Keep Teaching AI to Retrieve Information. What If We Taught It to Understand It Instead?

7
Comments
1 min read
Fine-Tuning vs RAG vs Prompt Engineering: Choosing the Right AI Strategy for Your Business

Fine-Tuning vs RAG vs Prompt Engineering: Choosing the Right AI Strategy for Your Business

Comments
4 min read
Why AI Hallucinates: Think of It Like a GPS with an Outdated Map

Why AI Hallucinates: Think of It Like a GPS with an Outdated Map

2
Comments
3 min read
KoutenDB v0.9.0: From a Locality Experiment to a Testable Database

KoutenDB v0.9.0: From a Locality Experiment to a Testable Database

Comments
5 min read
How I Built a Persona Chat Agent and Fought Hallucination — A RAG Story

How I Built a Persona Chat Agent and Fought Hallucination — A RAG Story

Comments
6 min read
Vidilearn: AI Knowledge Ingestion & Retrieval Gateway for LLMs, Agents, and MCP Servers

Vidilearn: AI Knowledge Ingestion & Retrieval Gateway for LLMs, Agents, and MCP Servers

2
Comments 2
1 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.