DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
My AI Agent Read 56 KB to Answer One Question. I Made It Stop.

My AI Agent Read 56 KB to Answer One Question. I Made It Stop.

Comments
4 min read
One Agent or Five? What I Learned Running a Team of AI Coders

One Agent or Five? What I Learned Running a Team of AI Coders

Comments
4 min read
Serving cheap when two models agree: a measured cost lever

Serving cheap when two models agree: a measured cost lever

3
Comments 2
5 min read
CAPE - Collaborative Agents Prompt Engineering

CAPE - Collaborative Agents Prompt Engineering

4
Comments
10 min read
Transformer as an Incomplete Cognitive Architecture: What It Captures Well and What It Misses (A11 Perspective)

Transformer as an Incomplete Cognitive Architecture: What It Captures Well and What It Misses (A11 Perspective)

Comments
4 min read
How I Cut LLM API Costs by 60% With 2 Lines of Code

How I Cut LLM API Costs by 60% With 2 Lines of Code

3
Comments
4 min read
ai, deepseek, machinelearning

ai, deepseek, machinelearning

1
Comments 2
5 min read
How NGOs Can Use LLMs Without Getting Burned

How NGOs Can Use LLMs Without Getting Burned

2
Comments 3
12 min read
webSLM: Fine-tuning, Compiling, and Running Domain-Specific Small Language Models Entirely in the Browser

webSLM: Fine-tuning, Compiling, and Running Domain-Specific Small Language Models Entirely in the Browser

2
Comments 1
15 min read
Stop Hardcoding Your Agent Workflows (or Don't): A Dev's Guide to Supervisor Delegation

Stop Hardcoding Your Agent Workflows (or Don't): A Dev's Guide to Supervisor Delegation

1
Comments
3 min read
We Obsessed Over Gateway Latency for a Month. Then We Looked at the Actual Numbers.

We Obsessed Over Gateway Latency for a Month. Then We Looked at the Actual Numbers.

2
Comments 1
4 min read
My LLM provider went down for 11 minutes. My code spent 4 of them in connect timeouts.

My LLM provider went down for 11 minutes. My code spent 4 of them in connect timeouts.

Comments
4 min read
agenttap: see exactly what your LLM SDK sent to the wire, with API keys scrubbed

agenttap: see exactly what your LLM SDK sent to the wire, with API keys scrubbed

Comments
4 min read
A six-concern production harness for Nemotron agents on Crusoe Managed Inference

A six-concern production harness for Nemotron agents on Crusoe Managed Inference

Comments
4 min read
llmfleet: pool many agents' turns into one Batch API call and save 50 percent

llmfleet: pool many agents' turns into one Batch API call and save 50 percent

Comments
4 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.