DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
LLM-as-a-Judge: I Built One From Scratch, Then Checked It Against Humans

LLM-as-a-Judge: I Built One From Scratch, Then Checked It Against Humans

Comments
4 min read
A Better LLM Judge? The Rubric Made My Small Model Worse

A Better LLM Judge? The Rubric Made My Small Model Worse

Comments
5 min read
How Modern Transformer Blocks Work — From RMSNorm to MoE

How Modern Transformer Blocks Work — From RMSNorm to MoE

Comments
5 min read
Active Working Memory: The RAM of Agentic Systems

Readers debate non-deterministic assembly

Active Working Memory: The RAM of Agentic Systems

10
Comments 8
5 min read
Why waiting longer makes voice AI worse

Why waiting longer makes voice AI worse

Comments 1
3 min read
Designing Edit Operations for AI Agents

Designing Edit Operations for AI Agents

1
Comments 2
5 min read
How to switch AI models without rewriting your app

How to switch AI models without rewriting your app

Comments
3 min read
I Built Byte Because OpenWebUI Kept Breaking

I Built Byte Because OpenWebUI Kept Breaking

5
Comments
1 min read
I gave my Cursor agent real tools without five API keys

I gave my Cursor agent real tools without five API keys

8
Comments 4
2 min read
From Transformer to ChatGPT: How One Paper Changed AI Engineering Forever

From Transformer to ChatGPT: How One Paper Changed AI Engineering Forever

1
Comments
3 min read
A Framework-Agnostic Testing Methodology for AI Agents (61 sources, 58 test blocks, OWASP Agentic Top 10)

A Framework-Agnostic Testing Methodology for AI Agents (61 sources, 58 test blocks, OWASP Agentic Top 10)

Comments
1 min read
Stop pasting your API keys into ChatGPT: a safer way to feed a codebase to an LLM

Stop pasting your API keys into ChatGPT: a safer way to feed a codebase to an LLM

Comments
2 min read
"LLM Inference Optimization: The Line Item That Decides If Your AI Ships"

"LLM Inference Optimization: The Line Item That Decides If Your AI Ships"

Comments
2 min read
Resurrecting Kepler: Getting Modern LLMs Running on a GTX 770 (Kernel 7.x)

Resurrecting Kepler: Getting Modern LLMs Running on a GTX 770 (Kernel 7.x)

1
Comments
4 min read
Caching LLM responses is just content addressing

Caching LLM responses is just content addressing

Comments
5 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.