DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
How I stopped my AI trip planner from inventing addresses

How I stopped my AI trip planner from inventing addresses

1
Comments
7 min read
Why AI Code Assistants Waste Context — and How RAG Fixes It

Why AI Code Assistants Waste Context — and How RAG Fixes It

Comments
18 min read
my hackathon submission

my hackathon submission

2
Comments
2 min read
Building a Healthcare AI Assistant with RAG, .NET, PostgreSQL, and pgvector

Building a Healthcare AI Assistant with RAG, .NET, PostgreSQL, and pgvector

Comments
3 min read
Stop Paying For Retrieval Latency On Chunks You Never Use In The Prompt

Stop Paying For Retrieval Latency On Chunks You Never Use In The Prompt

Comments
5 min read
# From Metadata to Knowledge Discovery: Why I Am Not Starting With a Chatbot

# From Metadata to Knowledge Discovery: Why I Am Not Starting With a Chatbot

Comments
3 min read
Claude Code 'Run Until Done' Mode, AI Concierge, & Mythos Scan for Curl Bugs

Claude Code 'Run Until Done' Mode, AI Concierge, & Mythos Scan for Curl Bugs

Comments
3 min read
There Is No Single "Best Model"

There Is No Single "Best Model"

Comments
2 min read
How I set up RAG evals in CI/CD so they actually catch regressions

How I set up RAG evals in CI/CD so they actually catch regressions

2
Comments
6 min read
Production Reranker Layer for RAG in Python: Cross-Encoder, Cohere Fallback, and Reciprocal Rank Fusion (Runnable Code)

Production Reranker Layer for RAG in Python: Cross-Encoder, Cohere Fallback, and Reciprocal Rank Fusion (Runnable Code)

Comments
10 min read
Beyond Vector Search: What RAG Actually Needs

Beyond Vector Search: What RAG Actually Needs

1
Comments
3 min read
We are Hiring! Senior GenAI Engineer — LLMs / RAG / Python (f/m/d) — Remote in Germany

We are Hiring! Senior GenAI Engineer — LLMs / RAG / Python (f/m/d) — Remote in Germany

Comments
2 min read
A hyped memory paper dropped. We didn't adopt it — and that was the point.

A hyped memory paper dropped. We didn't adopt it — and that was the point.

Comments
3 min read
Taming Unpredictable User Input: Building a RAG Triage Agent in Node.js

Taming Unpredictable User Input: Building a RAG Triage Agent in Node.js

2
Comments
2 min read
Vectorizing Real-Time Kafka Events

Vectorizing Real-Time Kafka Events

Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.