DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
RAG vs. Fine-Tuning: The AI Engineer's Decision Framework

RAG vs. Fine-Tuning: The AI Engineer's Decision Framework

4
Comments 1
3 min read
RAGBench: Stop Guessing, Start Measuring Your RAG Pipeline

RAGBench: Stop Guessing, Start Measuring Your RAG Pipeline

Comments 2
1 min read
Architecture of an Auditable Al Chatbot: Multi-Agent Routing with Open Knowledge Format

Architecture of an Auditable Al Chatbot: Multi-Agent Routing with Open Knowledge Format

Comments 2
9 min read
How MCP fetch servers keep getting SSRF wrong (and how I tried not to)

How MCP fetch servers keep getting SSRF wrong (and how I tried not to)

1
Comments
6 min read
Almanac's Company-Context Agent: How YC S26 Wires Internal Knowledge into Every LLM Call

Almanac's Company-Context Agent: How YC S26 Wires Internal Knowledge into Every LLM Call

Comments 1
6 min read
Securing Enterprise AI Pipelines: Using Azure Key Vault and AES Encryption for Sensitive Documents in Snowflake/Databricks AI Workloads

Securing Enterprise AI Pipelines: Using Azure Key Vault and AES Encryption for Sensitive Documents in Snowflake/Databricks AI Workloads

Comments 1
7 min read
BM25 Length Normalization: Why Long RAG Chunks Never Rank

BM25 Length Normalization: Why Long RAG Chunks Never Rank

Comments
7 min read
LlamaIndex makes RAG easy to build and hard to debug. Here is how I evaluate it.

LlamaIndex makes RAG easy to build and hard to debug. Here is how I evaluate it.

2
Comments
6 min read
Would your RAG eval suite notice if someone weakened the prompt?

Would your RAG eval suite notice if someone weakened the prompt?

7
Comments 7
3 min read
Firecrawl vs. Jina Reader vs. Tavily: Picking the Right Web-to-LLM Tool (and When You Need None of Them)

Firecrawl vs. Jina Reader vs. Tavily: Picking the Right Web-to-LLM Tool (and When You Need None of Them)

1
Comments
5 min read
483 tests passed, but Vestibule RAG framework wasn't installable — lessons from building with AI agents

483 tests passed, but Vestibule RAG framework wasn't installable — lessons from building with AI agents

1
Comments 5
3 min read
7 Reasons Your RAG Chatbot Fails in Production (And How to Fix Them in TypeScript)

7 Reasons Your RAG Chatbot Fails in Production (And How to Fix Them in TypeScript)

Comments 1
9 min read
A Developer's Checklist for Every RAG Lifecycle (Beyond Chunk-Embed-Search)

A Developer's Checklist for Every RAG Lifecycle (Beyond Chunk-Embed-Search)

Comments 1
2 min read
How Do You Actually Evaluate Your RAG App?

How Do You Actually Evaluate Your RAG App?

4
Comments 10
10 min read
Build Semantic Search for Legal Documents with Pinecone and GPT-4

Build Semantic Search for Legal Documents with Pinecone and GPT-4

Comments 1
12 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.