DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Four pgvector patterns that kept our RAG SaaS on one Postgres

Four pgvector patterns that kept our RAG SaaS on one Postgres

Comments
6 min read
Practical Agent Architecture: State, Failure Recovery, and the Hidden Variables of Reliable LLM Systems

Practical Agent Architecture: State, Failure Recovery, and the Hidden Variables of Reliable LLM Systems

1
Comments 1
5 min read
AI Agent Security, Open-Source Code Generation, and Frontier Models on Bedrock

AI Agent Security, Open-Source Code Generation, and Frontier Models on Bedrock

Comments
4 min read
Why Not Every AI Application Needs Vector Embeddings

Why Not Every AI Application Needs Vector Embeddings

1
Comments 1
7 min read
Production RAG at Scale: Lessons from Processing 10,000+ Listings Daily

Production RAG at Scale: Lessons from Processing 10,000+ Listings Daily

1
Comments 1
6 min read
RAG vs Fine-Tuning: Which Approach Should You Choose?

RAG vs Fine-Tuning: Which Approach Should You Choose?

Comments
3 min read
Building a multi-agent document-search copilot — Part 2: adaptive Hybrid, and a permission gate after the rank

Building a multi-agent document-search copilot — Part 2: adaptive Hybrid, and a permission gate after the rank

Comments 4
6 min read
RAG-Based Testing Series — Part 4: Edge Cases — What Breaks RAG & How to Catch It

RAG-Based Testing Series — Part 4: Edge Cases — What Breaks RAG & How to Catch It

7
Comments 1
14 min read
RAG-Based Testing Series — Part 3: Faithfulness & Hallucination Detection

RAG-Based Testing Series — Part 3: Faithfulness & Hallucination Detection

5
Comments
9 min read
I Contributed to Chroma's Open-Source Docs — Here's What I Changed and What I Learned

I Contributed to Chroma's Open-Source Docs — Here's What I Changed and What I Learned

Comments
4 min read
AI Output Provenance for SaaS: Trace Answers Before They Become Liability

AI Output Provenance for SaaS: Trace Answers Before They Become Liability

2
Comments
9 min read
RAG in 8 Layers: The Production Mental Model Most Tutorials Skip

RAG in 8 Layers: The Production Mental Model Most Tutorials Skip

Comments
16 min read
TradeMemory An AI-Powered Persistence Layer for Disciplined Trading

TradeMemory An AI-Powered Persistence Layer for Disciplined Trading

1
Comments
3 min read
The Data Ingestion Pipeline Nobody Designs Well Until Production Breaks It

The Data Ingestion Pipeline Nobody Designs Well Until Production Breaks It

Comments
5 min read
TradeMemory An AI-Powered Persistence Layer for Disciplined Trading

TradeMemory An AI-Powered Persistence Layer for Disciplined Trading

Comments
3 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.