DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
How NGOs Can Use LLMs Without Getting Burned

How NGOs Can Use LLMs Without Getting Burned

2
Comments 3
12 min read
We Obsessed Over Gateway Latency for a Month. Then We Looked at the Actual Numbers.

We Obsessed Over Gateway Latency for a Month. Then We Looked at the Actual Numbers.

2
Comments 1
4 min read
Stop Pasting Your Code Into ChatGPT For Debugging—Run LLMs Locally Instead

Stop Pasting Your Code Into ChatGPT For Debugging—Run LLMs Locally Instead

1
Comments
4 min read
A six-concern production harness for Nemotron agents on Crusoe Managed Inference

A six-concern production harness for Nemotron agents on Crusoe Managed Inference

Comments
4 min read
cachebench: stop finding out about prompt-cache regressions from the invoice

cachebench: stop finding out about prompt-cache regressions from the invoice

Comments
4 min read
I needed a stable cache key for LLM requests. The hard part was the input list order.

I needed a stable cache key for LLM requests. The hard part was the input list order.

Comments
4 min read
My LLM provider went down for 11 minutes. My code spent 4 of them in connect timeouts.

My LLM provider went down for 11 minutes. My code spent 4 of them in connect timeouts.

Comments
4 min read
agenttap: see exactly what your LLM SDK sent to the wire, with API keys scrubbed

agenttap: see exactly what your LLM SDK sent to the wire, with API keys scrubbed

Comments
4 min read
I shipped eight agent-stack repos in eight hours. Here's what made it possible.

I shipped eight agent-stack repos in eight hours. Here's what made it possible.

Comments
4 min read
llmfleet: pool many agents' turns into one Batch API call and save 50 percent

llmfleet: pool many agents' turns into one Batch API call and save 50 percent

Comments
4 min read
Compass v1.1.0 · we shipped a memory plugin that catches its own consumption drift

Compass v1.1.0 · we shipped a memory plugin that catches its own consumption drift

1
Comments
5 min read
We Let Sci-Fi Authors Code AI For Us

We Let Sci-Fi Authors Code AI For Us

2
Comments
4 min read
What 'Bring Your Own Model' (BYOK) Actually Means When You Adopt AI at Work

What 'Bring Your Own Model' (BYOK) Actually Means When You Adopt AI at Work

Comments
4 min read
.klickd v4.0.0 — Portable AI memory with constraints, strict schemas, and test vectors

.klickd v4.0.0 — Portable AI memory with constraints, strict schemas, and test vectors

Comments
6 min read
The Human in the Loop Doesn't Scale. I Kept Him Anyway.

The Human in the Loop Doesn't Scale. I Kept Him Anyway.

1
Comments 2
6 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.