DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
I gave an LLM 248 tools and accuracy dropped to 12%. Here's what fixed it.

I gave an LLM 248 tools and accuracy dropped to 12%. Here's what fixed it.

4
Comments
3 min read
Best Ways to Monitor Claude Code Token Usage and Costs in 2026

Best Ways to Monitor Claude Code Token Usage and Costs in 2026

2
Comments 7
6 min read
Teaching an AI to Play Dwarf Fortress: The Idea

Teaching an AI to Play Dwarf Fortress: The Idea

1
Comments 5
9 min read
I built a “deterministic” LLM text rephraser with a validation pipeline - looking for architectural feedback

I built a “deterministic” LLM text rephraser with a validation pipeline - looking for architectural feedback

Comments
3 min read
Why I Replaced LangChain with 15KB of httpx

Why I Replaced LangChain with 15KB of httpx

Comments
6 min read
Top AI Gateways with Semantic Caching and Dynamic Routing (2026 Guide)

Top AI Gateways with Semantic Caching and Dynamic Routing (2026 Guide)

1
Comments
7 min read
How to Design LLM Applications for Production: A System Design Guide

How to Design LLM Applications for Production: A System Design Guide

Comments
6 min read
I built memory decay for AI agents using the Ebbinghaus forgetting curve

I built memory decay for AI agents using the Ebbinghaus forgetting curve

24
Comments 2
2 min read
OpenTelemetry for LLM Applications: A Practical Guide with LaunchDarkly and Langfuse

OpenTelemetry for LLM Applications: A Practical Guide with LaunchDarkly and Langfuse

1
Comments 1
14 min read
Realtime steering: interrupt, barge-in, redirect, and guide the AI

Realtime steering: interrupt, barge-in, redirect, and guide the AI

1
Comments
4 min read
Building an Autonomous AI Agent for a Social Network: Lessons from the Chaos

Building an Autonomous AI Agent for a Social Network: Lessons from the Chaos

Comments
5 min read
Caching Strategies for LLM Systems (Part 3): Multi-Query Attention and Memory-Efficient Decoding

Caching Strategies for LLM Systems (Part 3): Multi-Query Attention and Memory-Efficient Decoding

Comments
5 min read
LocalAI QuickStart: Run OpenAI-Compatible LLMs Locally

LocalAI QuickStart: Run OpenAI-Compatible LLMs Locally

1
Comments
9 min read
Your Next.js Site Is Serving 26 KB of Noise to LLMs. Here's the Fix.

Your Next.js Site Is Serving 26 KB of Noise to LLMs. Here's the Fix.

1
Comments
3 min read
One command to add structured markup to your AI agent

One command to add structured markup to your AI agent

6
Comments 2
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.