DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Silent Model Swaps Are Eating Your LLM Budget — How to Detect Model Drift in Production

Silent Model Swaps Are Eating Your LLM Budget — How to Detect Model Drift in Production

1
Comments
4 min read
OpenClaw and Hermes agree on what an agent is. They disagree on what controls it.

OpenClaw and Hermes agree on what an agent is. They disagree on what controls it.

1
Comments
3 min read
Qwen 3.6 & llama.cpp Push Local Inference Limits on Consumer GPUs

Qwen 3.6 & llama.cpp Push Local Inference Limits on Consumer GPUs

Comments
3 min read
AI Weekly — 2026-05-15 to 2026-05-22 | The Agentic Inflection Is Real, But the Enterprise Gap Is Wider Than Ever

AI Weekly — 2026-05-15 to 2026-05-22 | The Agentic Inflection Is Real, But the Enterprise Gap Is Wider Than Ever

Comments
4 min read
Routing Event-Camera Pipelines Through an LLM Gateway: A Field Report

Routing Event-Camera Pipelines Through an LLM Gateway: A Field Report

Comments
4 min read
I tested cheap vs expensive LLMs across 3 real agent tasks. The cheap model won every time.

I tested cheap vs expensive LLMs across 3 real agent tasks. The cheap model won every time.

Comments
4 min read
Why Prompt Injection Won't Be "Fixed"

Why Prompt Injection Won't Be "Fixed"

1
Comments 3
9 min read
Measuring AI Gateway Failover: 30 Days of Production Data

Measuring AI Gateway Failover: 30 Days of Production Data

Comments
3 min read
Routing diffusion inference traffic across three providers

Routing diffusion inference traffic across three providers

Comments
4 min read
Evaluating LLM Output Quality In Production

Evaluating LLM Output Quality In Production

7
Comments 2
10 min read
Why RAG Isn't Enough: Building RationaleVault for Cognitive Continuity

Why RAG Isn't Enough: Building RationaleVault for Cognitive Continuity

Comments 1
4 min read
ToolRouter: Switch AI Coding Tools Freely Without Losing Context

ToolRouter: Switch AI Coding Tools Freely Without Losing Context

2
Comments
6 min read
Beyond the Stateless Prompt: Building an Auditable Product Intelligence Pipeline with Cascadeflow and Hindsight

Beyond the Stateless Prompt: Building an Auditable Product Intelligence Pipeline with Cascadeflow and Hindsight

Comments
5 min read
Putting an LLM Gateway in Front of Our Build Agents

Putting an LLM Gateway in Front of Our Build Agents

Comments
4 min read
Five ways your AI coding agent wastes tokens (and how to fix each one)

Five ways your AI coding agent wastes tokens (and how to fix each one)

2
Comments 1
6 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.