DEV Community

Testing

Find those bugs before your users do! 🐛

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Your test harness can silently not be testing the thing you think it is

Your test harness can silently not be testing the thing you think it is

Comments
4 min read
Shadow-Compare the Agent Patch. Merge Only Classified Divergences.

Shadow-Compare the Agent Patch. Merge Only Classified Divergences.

Comments
8 min read
My Bot's KPI: Trades 2 > Opportunities 1?! When Your Numerator and Denominator Don't Observe the Same Time.

My Bot's KPI: Trades 2 > Opportunities 1?! When Your Numerator and Denominator Don't Observe the Same Time.

Comments
4 min read
I said no data was leaving. On the first good run, two records left

I said no data was leaving. On the first good run, two records left

Comments
3 min read
Astro 7.3's Smallest New Feature Might Be the Most Useful One for Testing

Astro 7.3's Smallest New Feature Might Be the Most Useful One for Testing

Comments
2 min read
My own tool lied to me for three weeks: “declared” and “works” are different claims

My own tool lied to me for three weeks: “declared” and “works” are different claims

Comments
3 min read
Assumption Density Belongs Beside Pass Rate

Assumption Density Belongs Beside Pass Rate

Comments
6 min read
Stop Conditions for Free Inference in Agent Loops

Stop Conditions for Free Inference in Agent Loops

Comments
7 min read
The Assumption Gate: A Two-Hour AI Patch Lab

The Assumption Gate: A Two-Hour AI Patch Lab

Comments
6 min read
Stop Writing Ruby Tests in JS Template Literals: Meet rspec-wasm

Stop Writing Ruby Tests in JS Template Literals: Meet rspec-wasm

Comments
4 min read
Catch Tool Calls That Invent Missing Arguments

Catch Tool Calls That Invent Missing Arguments

Comments 1
7 min read
Zero Dependencies, Zero Regressions: What It Took to Build CLI-KDG from Scratch

Zero Dependencies, Zero Regressions: What It Took to Build CLI-KDG from Scratch

Comments
5 min read
Classify the Failure Before the Agent Edits a Test

Classify the Failure Before the Agent Edits a Test

Comments
7 min read
Fail-Open Defaults in Agent PRs: A Provenance Review

Fail-Open Defaults in Agent PRs: A Provenance Review

Comments
7 min read
A 10M-Token Budget Is a Test Plan: C++ Patch Trials on MonkeyCode's Free Tier

A 10M-Token Budget Is a Test Plan: C++ Patch Trials on MonkeyCode's Free Tier

Comments
5 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.