DEV Community

Running an AI Agent Fleet Series' Articles

Back to John's Series
Why I Rejected an Event Bus for My Solo Agent Fleet: State Is Truth, Events Are Rumors

Why I Rejected an Event Bus for My Solo Agent Fleet: State Is Truth, Events Are Rumors

Comments 7
7 min read
Stop Hooks as Hard Constraints: Enforcing Claude Code Behavior Outside the Model

Stop Hooks as Hard Constraints: Enforcing Claude Code Behavior Outside the Model

Comments 6
6 min read
No AI Claim Without a Kill Condition: Falsifier-Driven AI Decisions

No AI Claim Without a Kill Condition: Falsifier-Driven AI Decisions

1
Comments 2
4 min read
Stop AI Agent Drift Across Sessions With Versioned, Grep-able Rules

Stop AI Agent Drift Across Sessions With Versioned, Grep-able Rules

2
Comments 2
5 min read
Aggregating cron and launchd Into One Dashboard — Without Migrating Either

Aggregating cron and launchd Into One Dashboard — Without Migrating Either

Comments
6 min read
The 'You Decide' Reflex: Blocking AI-Agent Decision Punting with a Stop Hook

The 'You Decide' Reflex: Blocking AI-Agent Decision Punting with a Stop Hook

Comments 33
8 min read
Your AI Agent Folds When You Push Back: Measured Sycophancy and a Challenge-Triggered Verification Gate

Your AI Agent Folds When You Push Back: Measured Sycophancy and a Challenge-Triggered Verification Gate

Comments 25
8 min read
The Scary Metric Was Wrong, the Audit Still Paid: A 21-Agent Sweep of My Claude Code Fleet

The Scary Metric Was Wrong, the Audit Still Paid: A 21-Agent Sweep of My Claude Code Fleet

Comments
7 min read
Five Models, One Shared Blind Spot: What Multi-Model Fan-Out Catches and What It Can't

Five Models, One Shared Blind Spot: What Multi-Model Fan-Out Catches and What It Can't

Comments 2
6 min read
Verify the Output Surface: How 19 Green Tests Shipped Nine Broken Titles for Nine Days

Verify the Output Surface: How 19 Green Tests Shipped Nine Broken Titles for Nine Days

Comments
8 min read
Cross-Vendor Audit: What It Caught in My Own Model's Writing, and What It Got Wrong

Cross-Vendor Audit: What It Caught in My Own Model's Writing, and What It Got Wrong

Comments
6 min read
Green for Four Days While Nothing Shipped: A Reader Rebutted My Monitoring Post, and He Was Right

Green for Four Days While Nothing Shipped: A Reader Rebutted My Monitoring Post, and He Was Right

Comments 2
8 min read
I Asked Three Judges If I Was Wrong. One JSON Field Decided the Answer.

I Asked Three Judges If I Was Wrong. One JSON Field Decided the Answer.

Comments
6 min read
Your Agent Telemetry Ranks Your Routing Policy, Not Your Models

Your Agent Telemetry Ranks Your Routing Policy, Not Your Models

1
Comments 4
6 min read
My AI Subagent Faked the Verification Output I Asked It to Attach

My AI Subagent Faked the Verification Output I Asked It to Attach

2
Comments 14
6 min read
Deterministic Tool Adoption Gates: Score It, Don't Vibe It

Deterministic Tool Adoption Gates: Score It, Don't Vibe It

2
Comments 2
7 min read
Rotating the Hostile Seat: A Six-Round Adversarial Design Review Before Hardening an Agent

Rotating the Hostile Seat: A Six-Round Adversarial Design Review Before Hardening an Agent

Comments
7 min read
My Verification Gate Cleared on a Keyword, Not Evidence

Commenters expose a flaw in the gate logic

My Verification Gate Cleared on a Keyword, Not Evidence

11
Comments 22
6 min read
A Forced Dissent Slot Has a Floor: Read It by Convergence, Not Presence

A Forced Dissent Slot Has a Floor: Read It by Convergence, Not Presence

2
Comments
6 min read
Four Models Cited My Numbers Perfectly. One Still Misread Them.

Four Models Cited My Numbers Perfectly. One Still Misread Them.

Comments
6 min read
Our Quality Gate Was 24x Noisier Than What It Guarded

Our Quality Gate Was 24x Noisier Than What It Guarded

5
Comments 2
8 min read
Sub-Agent Metrics Are Not Comparable to Main-Thread Metrics

Sub-Agent Metrics Are Not Comparable to Main-Thread Metrics

1
Comments 12
8 min read