DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
OpenSquilla routes each turn to the cheapest model that can handle it

OpenSquilla routes each turn to the cheapest model that can handle it

1
Comments
3 min read
GLM 5.3: Zhipu's Open-Weight Model Excels at Coding and Cyber

GLM 5.3: Zhipu's Open-Weight Model Excels at Coding and Cyber

Comments 1
7 min read
Why Your AI Agent Forgets Everything Between Sessions (and the Three Honest Fixes)

Why Your AI Agent Forgets Everything Between Sessions (and the Three Honest Fixes)

1
Comments 2
4 min read
I Built an AI Agent With Tools. The First 50 Calls Were a Disaster.

I Built an AI Agent With Tools. The First 50 Calls Were a Disaster.

Comments
3 min read
Setting you up to fail

Setting you up to fail

Comments
3 min read
I Analyzed 4,788 AI Coding Sessions — Here's Where Your Tokens Actually Go

I Analyzed 4,788 AI Coding Sessions — Here's Where Your Tokens Actually Go

Comments
2 min read
MCP rug-pulls: how a "safe" AI tool turns malicious after you approve it

MCP rug-pulls: how a "safe" AI tool turns malicious after you approve it

3
Comments 2
3 min read
Answer Engine Optimization for Developers: What I Check First

Answer Engine Optimization for Developers: What I Check First

4
Comments 1
7 min read
A Line of Documentation Was Acting as a Global Config Flag

A Line of Documentation Was Acting as a Global Config Flag

Comments
6 min read
Self-hosting a lite agent backend on one TPU: Gemma 4 E2B + vLLM on a v5e-1

Self-hosting a lite agent backend on one TPU: Gemma 4 E2B + vLLM on a v5e-1

16
Comments 1
21 min read
Beyond The Single-Agent Ceiling: Scale Out With MCP Agent Teams

Beyond The Single-Agent Ceiling: Scale Out With MCP Agent Teams

2
Comments
17 min read
What Anthropic’s J-lens teaches us about debugging large language models

What Anthropic’s J-lens teaches us about debugging large language models

Comments
5 min read
RAG vs. Direct Context: I Tested Both on Real Documents, Here's What Broke

RAG vs. Direct Context: I Tested Both on Real Documents, Here's What Broke

11
Comments 1
5 min read
The LLM Waterfall Pattern: Never Let a Rate Limit Kill Your Workflow

The LLM Waterfall Pattern: Never Let a Rate Limit Kill Your Workflow

Comments
5 min read
What LLM Cost Calculators Get Wrong

What LLM Cost Calculators Get Wrong

3
Comments 11
8 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.