DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
We logged every rejected tool call for a month. A third were our validation being wrong, not the model.

We logged every rejected tool call for a month. A third were our validation being wrong, not the model.

1
Comments
2 min read
LLM Cost Optimization: How We Cut Reply Generation from $0.011 to $0.0009

LLM Cost Optimization: How We Cut Reply Generation from $0.011 to $0.0009

1
Comments
9 min read
The Model Context Protocol (MCP): what it is and how to build a server

The Model Context Protocol (MCP): what it is and how to build a server

Comments
7 min read
Designing Your Own AI Harness: A Deep Dive Into the Architecture of Agent Loops, Tools, Context, and Control

Designing Your Own AI Harness: A Deep Dive Into the Architecture of Agent Loops, Tools, Context, and Control

1
Comments 2
20 min read
Local MCP Development with Python and Kiro

Local MCP Development with Python and Kiro

Comments
8 min read
Beyond the Single Model: Why we Built an LLM Orchestrator with Consensus Engine?

Beyond the Single Model: Why we Built an LLM Orchestrator with Consensus Engine?

1
Comments
1 min read
Show HN: I replaced database deadlock victim selection with an LLM — here's the data

Show HN: I replaced database deadlock victim selection with an LLM — here's the data

Comments
4 min read
Kronos Financial LLM, Local AI Health Checks & Code-RAG Benchmarking Insights

Kronos Financial LLM, Local AI Health Checks & Code-RAG Benchmarking Insights

Comments
3 min read
Top 5 Open Source MCP Gateways in 2026, Compared

Top 5 Open Source MCP Gateways in 2026, Compared

15
Comments
5 min read
Does Quantization Break Tool-Calling? I Measured It on a 4GB Laptop GPU (BFCL, 3 Seeds, Bootstrap 95% CI)

Model family outperforms size as a predictor

Does Quantization Break Tool-Calling? I Measured It on a 4GB Laptop GPU (BFCL, 3 Seeds, Bootstrap 95% CI)

3
Comments 10
4 min read
Deterministic Data Engineering With AI Harnesses: Using Claude Code, Codex, Antigravity, and OpenCode for Data Work You Can Actually Trust

Deterministic Data Engineering With AI Harnesses: Using Claude Code, Codex, Antigravity, and OpenCode for Data Work You Can Actually Trust

Comments
19 min read
Gates Earned From Failure: a cost test for agent guardrails

Gates Earned From Failure: a cost test for agent guardrails

Comments
16 min read
Why most AI apps fail in production (not in demos)

Why most AI apps fail in production (not in demos)

Comments
1 min read
The --schema-only flag that makes enterprise customers comfortable with AI

The --schema-only flag that makes enterprise customers comfortable with AI

Comments
4 min read
Shipping 100,000 construction PDFs a month: what actually breaks

Shipping 100,000 construction PDFs a month: what actually breaks

1
Comments
17 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.