DEV Community

#agents

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
The Looping Principle: A Simple Mental Model for Understanding AI Agents

The Looping Principle: A Simple Mental Model for Understanding AI Agents

Comments
4 min read
When a human edits an agent's output, where does that decision go?

When a human edits an agent's output, where does that decision go?

1
Comments
2 min read
What actually happened when my agent worked on Frantic

What actually happened when my agent worked on Frantic

1
Comments
2 min read
The System Learns to Read Without Obeying

The System Learns to Read Without Obeying

1
Comments
5 min read
I Looked for Where My Publishing Agent Needed a State Machine. The State Machine Was Already Live on dev.to.

I Looked for Where My Publishing Agent Needed a State Machine. The State Machine Was Already Live on dev.to.

Comments 1
4 min read
Karpathy: Agent Performance Gap Is in the Harness, Not the Model

Karpathy: Agent Performance Gap Is in the Harness, Not the Model

Comments 1
2 min read
Our AI Agent Failed 5 Times in One Day. Here is Why It Never Happened Again.

Our AI Agent Failed 5 Times in One Day. Here is Why It Never Happened Again.

1
Comments 6
2 min read
完美的平庸

完美的平庸

Comments
1 min read
enable AI's Full Potential: Structure Your Codebase for Agent Success

enable AI's Full Potential: Structure Your Codebase for Agent Success

Comments 1
6 min read
How to Build a Production Agent Harness

How to Build a Production Agent Harness

24
Comments 1
11 min read
MCP: A Complete Guide from Zero to Maximum, from Tools to Cross-Regional Discovery with Cryptographic Trust Scoring.

MCP: A Complete Guide from Zero to Maximum, from Tools to Cross-Regional Discovery with Cryptographic Trust Scoring.

4
Comments
8 min read
China Published Its First AI Agent Trust Standard. We Mapped It to 2.3 Million Real Agents.

China Published Its First AI Agent Trust Standard. We Mapped It to 2.3 Million Real Agents.

Comments
8 min read
Running a 122B Parameter Model and Agent Locally on AMD MI300X GPU — What I Learned

Running a 122B Parameter Model and Agent Locally on AMD MI300X GPU — What I Learned

1
Comments
7 min read
Why Agent Evaluation Is Harder Than Model Evaluation

Builder scars reveal untrustworthy agent paths

Why Agent Evaluation Is Harder Than Model Evaluation

19
Comments 25
8 min read
Why your agent benchmarks are lying to you

Why your agent benchmarks are lying to you

1
Comments 1
2 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.