DEV Community

#dailybuild2026

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Two agents, two companies, one protocol: building a visible agent-to-agent negotiation

Two agents, two companies, one protocol: building a visible agent-to-agent negotiation

Comments 1
11 min read
Building a Rubik's Cube Decision Arena with 6 Decision Models (Laya, Clef, GLiNER 2.5, Kev and Strands Decider)

Building a Rubik's Cube Decision Arena with 6 Decision Models (Laya, Clef, GLiNER 2.5, Kev and Strands Decider)

Comments
14 min read
Building Anvil: code-as-action with a capability sandbox that explains its refusals.

Building Anvil: code-as-action with a capability sandbox that explains its refusals.

1
Comments
8 min read
How Arbiter works, why it was built, and the four measurements it made about an LLM judge that I had to look at twice.

How Arbiter works, why it was built, and the four measurements it made about an LLM judge that I had to look at twice.

Comments
12 min read
How I built a time-travel debugger for LLM agents on Burr, and what "replay the unchanged branch and prove nothing changed" actually takes.

How I built a time-travel debugger for LLM agents on Burr, and what "replay the unchanged branch and prove nothing changed" actually takes.

Comments
8 min read
Guardrails are a lie until they're arithmetic

Guardrails are a lie until they're arithmetic

1
Comments 1
11 min read
Your agent is a long-running process with side effects. Kill it and watch what happens.

Your agent is a long-running process with side effects. Kill it and watch what happens.

Comments
11 min read
Sleeptime memory consolidation: an agent that edits its own memory and proves it still remembers

Sleeptime memory consolidation: an agent that edits its own memory and proves it still remembers

Comments
15 min read
Luthier: an agent that audits an AI coding agent's harness against the repo it governs.

Luthier: an agent that audits an AI coding agent's harness against the repo it governs.

Comments
9 min read
One UI, Any Agent: Proving Framework Independence with AG-UI, Mastra, and React

One UI, Any Agent: Proving Framework Independence with AG-UI, Mastra, and React

Comments
5 min read
Bloom: I Made Two LLMs Paint the Same Sentence and Measured What Happened

Bloom: I Made Two LLMs Paint the Same Sentence and Measured What Happened

Comments
7 min read
Glyph: I made two AI models fight inside real terminals and the only honest judge is a PTY

Glyph: I made two AI models fight inside real terminals and the only honest judge is a PTY

1
Comments
4 min read
Building The Confidence Curve

Building The Confidence Curve

1
Comments
14 min read
Automating LLM A/B Testing: How to Build a Cross-Model Evaluation Harness Programmatically

Automating LLM A/B Testing: How to Build a Cross-Model Evaluation Harness Programmatically

Comments
6 min read
How I built Refract: a live arena where two LLMs fight in GLSL

How I built Refract: a live arena where two LLMs fight in GLSL

1
Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.