DEV Community

#rag

Retrieval augmented generation, or RAG, is an architectural approach that can improve the efficacy of large language model (LLM) applications by leveraging custom data.

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
# Building a Personal Notes Assistant with RAG, Amazon Bedrock, and Pinecone

# Building a Personal Notes Assistant with RAG, Amazon Bedrock, and Pinecone

2
Comments
5 min read
About Best in IT: Practical AI, Automation and Developer Tools

About Best in IT: Practical AI, Automation and Developer Tools

1
Comments
3 min read
Building a RAG Retrieval Service: pgvector, Embedding Migrations, and Provenance Tracking

Building a RAG Retrieval Service: pgvector, Embedding Migrations, and Provenance Tracking

Comments
3 min read
LlamaIndex re-retrieves your chunks — and re-sends up to 75% of your context — on every chat turn

LlamaIndex re-retrieves your chunks — and re-sends up to 75% of your context — on every chat turn

Comments
3 min read
Your Documents, Chunked and Searchable: The Knowledge Base in ByteChef

Your Documents, Chunked and Searchable: The Knowledge Base in ByteChef

Comments
7 min read
Building a Closed-Domain Agentic AI Knowledge Assistant with Hybrid RAG

Building a Closed-Domain Agentic AI Knowledge Assistant with Hybrid RAG

2
Comments
2 min read
Enterprise RAG Without Per-Token Pricing: DeepSeek R1 on SageMaker with OpenSearch

Enterprise RAG Without Per-Token Pricing: DeepSeek R1 on SageMaker with OpenSearch

1
Comments
5 min read
Build a Local RAG Chatbot for Trading Research Using Ollama + Termux (Zero API Cost)

Build a Local RAG Chatbot for Trading Research Using Ollama + Termux (Zero API Cost)

Comments
6 min read
Securing a RAG Pipeline — The Threats I Designed Against and the Ones I Didn't

Securing a RAG Pipeline — The Threats I Designed Against and the Ones I Didn't

1
Comments 1
8 min read
My notes became a personal RAG with no embeddings

My notes became a personal RAG with no embeddings

Comments 2
8 min read
Local Embeddings vs. API Embeddings — Why I Chose sentence-transformers

Local Embeddings vs. API Embeddings — Why I Chose sentence-transformers

Comments 1
6 min read
Multilingual RAG Architecture That Works in Production

Multilingual RAG Architecture That Works in Production

Comments 1
7 min read
Model Cascade: making LLM classification cheaper

Model Cascade: making LLM classification cheaper

Comments
5 min read
Mind Discipline: Why Our AI Advisor Only Reads Hand-Crafted Contracts

Mind Discipline: Why Our AI Advisor Only Reads Hand-Crafted Contracts

1
Comments
9 min read
从 Demo 到生产:那些真正让 AI Agent 敢上线的护栏

从 Demo 到生产:那些真正让 AI Agent 敢上线的护栏

Comments
2 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.