DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Redefining the Role of Google Apps Script in the Era of Generative AI

Redefining the Role of Google Apps Script in the Era of Generative AI

8
Comments
47 min read
Why I run speech-to-text locally instead of calling a cloud API

Why I run speech-to-text locally instead of calling a cloud API

4
Comments 2
3 min read
"Your cache hit rate is low" — true, and worth $0.16

"Your cache hit rate is low" — true, and worth $0.16

3
Comments 14
7 min read
I watched my LLM bill for 30 days. The 30x cache lever is real.

I watched my LLM bill for 30 days. The 30x cache lever is real.

1
Comments 1
3 min read
Building an Observable AI Market Research Agent with SigNoz

Building an Observable AI Market Research Agent with SigNoz

Comments
1 min read
Loop Engineering Is Mostly Papering Over a Model That Won't Converge

Loop Engineering Is Mostly Papering Over a Model That Won't Converge

2
Comments 2
4 min read
I Trained a 6.4M-Parameter Transformer From Scratch to Talk About Recipes

I Trained a 6.4M-Parameter Transformer From Scratch to Talk About Recipes

1
Comments
5 min read
Empire LLM for Codex: AI Code Review Without the Chaos

Empire LLM for Codex: AI Code Review Without the Chaos

Comments
5 min read
Is Speculative Decoding's Speedup a Hardware Problem or a Model Problem?

Is Speculative Decoding's Speedup a Hardware Problem or a Model Problem?

Comments
8 min read
I built a CLI that tells you if your codebase fits an LLM's context window

I built a CLI that tells you if your codebase fits an LLM's context window

4
Comments
2 min read
Spec Suite: Putting an End to Hallucinated AI Documentation

Spec Suite: Putting an End to Hallucinated AI Documentation

Comments
4 min read
I built a production AI agent as a Honda service advisor. Then I read the textbook.

I built a production AI agent as a Honda service advisor. Then I read the textbook.

Comments
7 min read
CacheGuard

CacheGuard

Comments
6 min read
Claude Opus 5 leads on agentic work — and undercuts Fable 5 on cost

Claude Opus 5 leads on agentic work — and undercuts Fable 5 on cost

Comments
2 min read
I tried to build a "token optimization stack" for coding agents. Here's why I killed it.

A 97% savings metric masked silent failures

I tried to build a "token optimization stack" for coding agents. Here's why I killed it.

5
Comments 10
7 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.