DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
The Multi-Runtime Agent Problem: Why Your Team Needs More Than One Runtime

The Multi-Runtime Agent Problem: Why Your Team Needs More Than One Runtime

Comments
5 min read
3 Assumptions That Broke Before I Got gpt-oss-120b Working Through the Anthropic SDK

3 Assumptions That Broke Before I Got gpt-oss-120b Working Through the Anthropic SDK

1
Comments
6 min read
There was no independent, measured view of AI-API latency by region — so I built one

There was no independent, measured view of AI-API latency by region — so I built one

Comments
4 min read
Your LLM gateway is throwing away the data that would improve your prompts

Your LLM gateway is throwing away the data that would improve your prompts

2
Comments 1
4 min read
Gemma 2's Architecture: More Performance from Less Model

Gemma 2's Architecture: More Performance from Less Model

Comments
3 min read
Why Most RAG Systems Fail in Production: The Hidden Architecture Problems Behind AI Search

Why Most RAG Systems Fail in Production: The Hidden Architecture Problems Behind AI Search

3
Comments 5
12 min read
Langfuse alternatives: 6 LLM observability tools, sorted by the thing that bites you in month eight

Langfuse alternatives: 6 LLM observability tools, sorted by the thing that bites you in month eight

Comments
4 min read
I Cut My AI Agent's Token Bill by 62% in One Weekend. Here's the Receipts.

I Cut My AI Agent's Token Bill by 62% in One Weekend. Here's the Receipts.

Comments
5 min read
Designing a Self-Prompting Agent Harness with Per-Task Prompt, Tool, and Strategy Synthesis

Designing a Self-Prompting Agent Harness with Per-Task Prompt, Tool, and Strategy Synthesis

Comments
17 min read
Fix LLM formatting in the tool layer, not the prompt

Fix LLM formatting in the tool layer, not the prompt

1
Comments 1
8 min read
Your RAG Retrieved the Right Documents but Still Gave the Wrong Answer

Your RAG Retrieved the Right Documents but Still Gave the Wrong Answer

Comments
2 min read
Your LLM reads the whole file. It doesn't have to.

Your LLM reads the whole file. It doesn't have to.

Comments
4 min read
How to Stop AI Agent Cost Blowups Before They Happen

How to Stop AI Agent Cost Blowups Before They Happen

1
Comments 2
5 min read
Unit Test AI Guide — Zero Hallucination, Cross-Stack Standard

Unit Test AI Guide — Zero Hallucination, Cross-Stack Standard

Comments
11 min read
Take your benchmark to the people who can kill it

Take your benchmark to the people who can kill it

Comments
9 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.