DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
"Smart Model Routing: Why Your AI Agent Shouldn't Use the Same Model for Everything"

"Smart Model Routing: Why Your AI Agent Shouldn't Use the Same Model for Everything"

Comments
2 min read
CAG: The Simpler Way to Ground Your LLM

CAG: The Simpler Way to Ground Your LLM

Comments
3 min read
Mem0 vs Letta vs Zep: Which Should You Use for Agent Memory?

Mem0 vs Letta vs Zep: Which Should You Use for Agent Memory?

2
Comments 1
9 min read
Evaluating Large Language Models: The Pitfall of Overfitting in RAG

Evaluating Large Language Models: The Pitfall of Overfitting in RAG

Comments
2 min read
Idle Drift

Idle Drift

Comments
4 min read
Evaluating Large Language Models: The Overfitting Problem

Evaluating Large Language Models: The Overfitting Problem

Comments
2 min read
Air-gapped code review with Ollama: when the diff never leaves the machine

Air-gapped code review with Ollama: when the diff never leaves the machine

Comments
4 min read
# Building AIRAG Jobs: A Job Board for LLM, RAG & AI Agent Engineers

# Building AIRAG Jobs: A Job Board for LLM, RAG & AI Agent Engineers

Comments
2 min read
Stop Scattering LLM Code Across Your Codebase

Stop Scattering LLM Code Across Your Codebase

1
Comments
2 min read
Building a RAG System from Scratch — Wrap-up and What Comes Next

Building a RAG System from Scratch — Wrap-up and What Comes Next

Comments
4 min read
Building a RAG System from Scratch — Design Decisions Explained

Building a RAG System from Scratch — Design Decisions Explained

Comments
4 min read
How to Run Reliable Local LLM Agents on an RTX 3090: A Benchmark (5 Models, Priced in Watts)

How to Run Reliable Local LLM Agents on an RTX 3090: A Benchmark (5 Models, Priced in Watts)

1
Comments
3 min read
Introducing Hearth v0.3.0: Scale Long-Tail LLMs to Zero on Private Kubernetes

Introducing Hearth v0.3.0: Scale Long-Tail LLMs to Zero on Private Kubernetes

2
Comments 3
5 min read
Duo Pipeline: Cutting AI Agent Costs by 70% with Adaptive Routing

Duo Pipeline: Cutting AI Agent Costs by 70% with Adaptive Routing

Comments
2 min read
Building a RAG System from Scratch — Tool Use: Let the LLM Search Autonomously

Building a RAG System from Scratch — Tool Use: Let the LLM Search Autonomously

Comments
5 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.