DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
The Open-Model Cost Chart Everyone's Sharing Is API Prices. Here's What Self-Hosting Actually Gets You (Measured)

The Open-Model Cost Chart Everyone's Sharing Is API Prices. Here's What Self-Hosting Actually Gets You (Measured)

1
Comments
5 min read
"GLM-5.2 and the Open-Weight Tipping Point"

"GLM-5.2 and the Open-Weight Tipping Point"

Comments 1
2 min read
Context Rot: Why Your AI Coding Agent Gets Dumber Mid-Session (and How I Stopped It)

Context Rot: Why Your AI Coding Agent Gets Dumber Mid-Session (and How I Stopped It)

Comments
4 min read
A diagram is data, not a drawing

A diagram is data, not a drawing

2
Comments 1
5 min read
Context rot is real. You can compile it away.

Context rot is real. You can compile it away.

1
Comments
2 min read
The Orchestration Bottleneck: Why Your Agent Infrastructure Needs Two Layers in 2026

The Orchestration Bottleneck: Why Your Agent Infrastructure Needs Two Layers in 2026

Comments
5 min read
Harvesting a regression test set from gateway logs with a plugin

Harvesting a regression test set from gateway logs with a plugin

Comments
4 min read
Use a flat-priced, auto-routing LLM API in Aider or Cline — one npx command

Use a flat-priced, auto-routing LLM API in Aider or Cline — one npx command

1
Comments 1
2 min read
Unifying image inputs across three vision providers behind Bifrost

Unifying image inputs across three vision providers behind Bifrost

Comments
4 min read
Claude Code retries rate-limit errors for API keys, not for your Max plan

Claude Code retries rate-limit errors for API keys, not for your Max plan

Comments
4 min read
Coding Senza Compiacenza: Come Far Dire "No" agli Agenti IA

Coding Senza Compiacenza: Come Far Dire "No" agli Agenti IA

7
Comments
8 min read
Designing a Synthetic Data Pipeline for Persian LLM Fine Tuning: From Topic Graphs to QLoRA Evaluation

Designing a Synthetic Data Pipeline for Persian LLM Fine Tuning: From Topic Graphs to QLoRA Evaluation

Comments
4 min read
I gave myself an AI advisory board — three models argue, I decide

I gave myself an AI advisory board — three models argue, I decide

Comments
3 min read
Semantic caching our flaky-test summariser: 58% fewer LLM calls

Semantic caching our flaky-test summariser: 58% fewer LLM calls

Comments
4 min read
Route Every Prompt to the Cheapest Model: Building a Multi-LLM Cost Optimizer with Pydantic AI

Route Every Prompt to the Cheapest Model: Building a Multi-LLM Cost Optimizer with Pydantic AI

Comments
6 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.