DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Building a Document-RAG Agent on GCP's Agent Development Kit (ADK)

Building a Document-RAG Agent on GCP's Agent Development Kit (ADK)

2
Comments
14 min read
Building RAG that doesn't hallucinate

Building RAG that doesn't hallucinate

2
Comments
4 min read
Where Does RAG Actually Cost You Money? I Decided to Stop Guessing.

Debunks the myth that embeddings drive up bills

Where Does RAG Actually Cost You Money? I Decided to Stop Guessing.

10
Comments 4
4 min read
Why Organizations Forget Even When Nothing Is Deleted

Why Organizations Forget Even When Nothing Is Deleted

5
Comments 22
2 min read
I Was Optimizing Ranking While the Real Problem Was Selection

I Was Optimizing Ranking While the Real Problem Was Selection

3
Comments 4
2 min read
Qdrant vs Pinecone: Self-Hosted Vector Search for Production RAG

Qdrant vs Pinecone: Self-Hosted Vector Search for Production RAG

Comments 3
11 min read
How Japan’s Research Labs Are Building RAG Systems That Actually Work — And What Western Teams Keep Getting Wrong

How Japan’s Research Labs Are Building RAG Systems That Actually Work — And What Western Teams Keep Getting Wrong

1
Comments
4 min read
Replacing Cross-Encoder Reranking with a Weighted Hybrid Score

Replacing Cross-Encoder Reranking with a Weighted Hybrid Score

Comments
5 min read
Stop Paying for Embedding APIs: Local Hybrid Search with Hippo

Stop Paying for Embedding APIs: Local Hybrid Search with Hippo

1
Comments 4
7 min read
1st post

1st post

Comments 1
1 min read
AI Agent Orchestration: Proxmox Automation, OpenAI Data Agents & Azure Serverless Runtime

AI Agent Orchestration: Proxmox Automation, OpenAI Data Agents & Azure Serverless Runtime

Comments
3 min read
I Built a Simple RAG App with LangChain, OpenAI, and Pinecone

I Built a Simple RAG App with LangChain, OpenAI, and Pinecone

2
Comments 1
10 min read
Phase 4: Retrieval Quality & Grounded Answers

Why closest matches aren't always relevant

Phase 4: Retrieval Quality & Grounded Answers

6
Comments 10
12 min read
Why Most RAG Systems Fail in Production: The Hidden Architecture Problems Behind AI Search

Why Most RAG Systems Fail in Production: The Hidden Architecture Problems Behind AI Search

3
Comments 5
12 min read
Your RAG Retrieved the Right Documents but Still Gave the Wrong Answer

Your RAG Retrieved the Right Documents but Still Gave the Wrong Answer

Comments
2 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.