DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Building a RAG Pipeline That Stays Fresh with Live Web Data

Building a RAG Pipeline That Stays Fresh with Live Web Data

Comments
5 min read
Claude 4.5 Speculative Tooling: Why Your Java Backend Needs Idempotency or It’s Dead

Claude 4.5 Speculative Tooling: Why Your Java Backend Needs Idempotency or It’s Dead

Comments
2 min read
50% Compliance, Not 0%: How a Logging Spike Almost Triggered the Wrong Architecture Rewrite

50% Compliance, Not 0%: How a Logging Spike Almost Triggered the Wrong Architecture Rewrite

Comments
1 min read
Compass v1.1.0 · we shipped a memory plugin that catches its own consumption drift

Compass v1.1.0 · we shipped a memory plugin that catches its own consumption drift

Comments 1
5 min read
Tool-use API design for LLMs: 5 patterns that prevent agent loops and silent failures

Tool-use API design for LLMs: 5 patterns that prevent agent loops and silent failures

Comments
9 min read
Stop Preloading Everything: How We Cut AI Agent Context by 50–87% with Lazy Discovery

Stop Preloading Everything: How We Cut AI Agent Context by 50–87% with Lazy Discovery

Comments
9 min read
A Picture Is Worth Ten Thousand Tokens

A Picture Is Worth Ten Thousand Tokens

Comments
6 min read
The Ultimate Developer's Directory: 282+ AI Tools & Agents You Need to Try

The Ultimate Developer's Directory: 282+ AI Tools & Agents You Need to Try

Comments
9 min read
ElevenLabs Conversational AI survey bot — reducing latency and robotic feel, plus initial delay issue

ElevenLabs Conversational AI survey bot — reducing latency and robotic feel, plus initial delay issue

Comments
1 min read
Taming multi-invoice PDFs and building a customer dashboard

Taming multi-invoice PDFs and building a customer dashboard

Comments
2 min read
LLM Study Diary #2: Tokenization

LLM Study Diary #2: Tokenization

Comments
2 min read
You Vibe-Coded Your SaaS Landing Page — Google Can't See It

You Vibe-Coded Your SaaS Landing Page — Google Can't See It

Comments
2 min read
LLM Foundry on a tiny model: the stack still does the heavy lifting

LLM Foundry on a tiny model: the stack still does the heavy lifting

Comments
1 min read
llama.cpp MTP Beta, Gemma GGUF Fixes, & Sentinel Local-First AI Coding App

llama.cpp MTP Beta, Gemma GGUF Fixes, & Sentinel Local-First AI Coding App

Comments
3 min read
Why Dense Search Fails in Production RAG — And How Hybrid Search Fixes It

Why Dense Search Fails in Production RAG — And How Hybrid Search Fixes It

1
Comments 3
5 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.