DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
The Biggest Pitfall in GraphRAG: One Entity, Seven Identities

The Biggest Pitfall in GraphRAG: One Entity, Seven Identities

Comments 1
7 min read
A $0.25 model beat a $3 model -- with better context

A $0.25 model beat a $3 model -- with better context

Comments
7 min read
Migrating vector embeddings in production without downtime

Migrating vector embeddings in production without downtime

30
Comments 2
6 min read
Speaking the Corpus's Language: How Multilingual RAG Stays Coherent Across Turns

Speaking the Corpus's Language: How Multilingual RAG Stays Coherent Across Turns

Comments 1
8 min read
10 Chunking Strategies That Make or Break Your RAG Pipeline

10 Chunking Strategies That Make or Break Your RAG Pipeline

18
Comments
7 min read
How I Built a RAG-Based Law Assistant with LangChain and FAISS

How I Built a RAG-Based Law Assistant with LangChain and FAISS

2
Comments 2
3 min read
Integrating AI into a Legacy Broadcasting CMS(Content-uploading Manager System): Architecture Design

Integrating AI into a Legacy Broadcasting CMS(Content-uploading Manager System): Architecture Design

Comments
5 min read
Our Group project was chaos until this agent

Our Group project was chaos until this agent

Comments 1
6 min read
Stop Losing Your Health Data! Build a Lifelong Electronic Health Record (EHR) System with Neo4j and GraphRAG đŸ„đŸ’»

Stop Losing Your Health Data! Build a Lifelong Electronic Health Record (EHR) System with Neo4j and GraphRAG đŸ„đŸ’»

1
Comments
3 min read
Building a Scalable RAG Backend with Cloud Run Jobs and AlloyDB

Building a Scalable RAG Backend with Cloud Run Jobs and AlloyDB

31
Comments 4
6 min read
Bringing The Receipts - 95% AI LLM Token Savings

Bringing The Receipts - 95% AI LLM Token Savings

1
Comments
10 min read
Building a Perplexity Clone for Local LLMs in 50 Lines of Python

Building a Perplexity Clone for Local LLMs in 50 Lines of Python

1
Comments 1
6 min read
Scaling LLMs at the Edge: A journey through distillation, routers, and embeddings

Scaling LLMs at the Edge: A journey through distillation, routers, and embeddings

1
Comments
20 min read
15 Engineering Decisions Behind RAG Hybrid Search

15 Engineering Decisions Behind RAG Hybrid Search

12
Comments
9 min read
RAG + FastAPI in Action: Creating a Smart Business Analytics Dashboard in Python

RAG + FastAPI in Action: Creating a Smart Business Analytics Dashboard in Python

Comments
9 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.