DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Retrieval-Augmented Self-Recall — Part 5: The Gap Threshold That Didn't Transfer

Retrieval-Augmented Self-Recall — Part 5: The Gap Threshold That Didn't Transfer

1
Comments
4 min read
Your RAG System Is Lying To You About That Table

Your RAG System Is Lying To You About That Table

13
Comments 2
4 min read
A Chinese 8B model beat the Western 8B models at Japanese RAG. I still wouldn't put it in the default deployment — and that distinction is the point.

A Chinese 8B model beat the Western 8B models at Japanese RAG. I still wouldn't put it in the default deployment — and that distinction is the point.

Comments
4 min read
RAG should never be your default

RAG should never be your default

Comments
3 min read
Building a Grounded RAG Assistant: Why Citation Enforcement Matters More Than Retrieval

Building a Grounded RAG Assistant: Why Citation Enforcement Matters More Than Retrieval

Comments 2
1 min read
AI Agents Level Up Workflows: Terraform MCP, WebMCP, Pinecone Integrations

AI Agents Level Up Workflows: Terraform MCP, WebMCP, Pinecone Integrations

Comments
4 min read
The Source-of-Truth Problem Every Enterprise AI Team Faces

The Source-of-Truth Problem Every Enterprise AI Team Faces

1
Comments
3 min read
RAG is for finding. Full context is for deciding.

RAG is for finding. Full context is for deciding.

4
Comments
4 min read
Why my first RAG layer starts in Postgres, not in a standalone vector database

Why my first RAG layer starts in Postgres, not in a standalone vector database

Comments
3 min read
Moonshot AI's Kimi K3 Is Here: A 2.8 Trillion Parameter Open MoE Model That Pushes Long-Context AI Forward

Moonshot AI's Kimi K3 Is Here: A 2.8 Trillion Parameter Open MoE Model That Pushes Long-Context AI Forward

1
Comments
3 min read
Beyond Chatbots: Wrapping My RAG Agent in an MCP Server

Beyond Chatbots: Wrapping My RAG Agent in an MCP Server

1
Comments 5
2 min read
Agent RAG in Solon: Make Retrieval a Tool, Not a One-Shot Pre-Step

Agent RAG in Solon: Make Retrieval a Tool, Not a One-Shot Pre-Step

Comments 5
5 min read
I built a "boring" RAG demo over World Cup data — SQLite, sqlite-vec, and no framework

I built a "boring" RAG demo over World Cup data — SQLite, sqlite-vec, and no framework

Comments
4 min read
Token Cost Optimization: How to Cut LLM Inference Spend Without Cutting Quality

Token Cost Optimization: How to Cut LLM Inference Spend Without Cutting Quality

1
Comments
5 min read
Your RAG Stack Is Solving the 2023 Problem

Your RAG Stack Is Solving the 2023 Problem

8
Comments 1
7 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.