DEV Community

Testing

Find those bugs before your users do! 🐛

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
From Single Files to Scenario Suites: Batch Validation in the OWASP Agent Security Regression Harness

From Single Files to Scenario Suites: Batch Validation in the OWASP Agent Security Regression Harness

2
Comments
3 min read
I spent hours writing unit tests – so I made an LLM do it (and learned what not to do)

I spent hours writing unit tests – so I made an LLM do it (and learned what not to do)

Comments
4 min read
I ran my own reliability tool on my production system. 90% of the findings were wrong — and that was the most valuable bug I've fixed.

I ran my own reliability tool on my production system. 90% of the findings were wrong — and that was the most valuable bug I've fixed.

1
Comments 1
6 min read
LLM-as-judge disagrees with itself between runs

LLM-as-judge disagrees with itself between runs

4
Comments 2
4 min read
Selenium & Python: The Complete Guide to Web Automation Testing

Selenium & Python: The Complete Guide to Web Automation Testing

Comments
6 min read
How I Built 5-Layer AI Quality Architecture Across 5 Production AI Systems

How I Built 5-Layer AI Quality Architecture Across 5 Production AI Systems

1
Comments
7 min read
設計你自己的 Multi-AI Coding Pipeline:一份可搬走的參考架構

設計你自己的 Multi-AI Coding Pipeline:一份可搬走的參考架構

Comments
1 min read
Design Your Own Multi-AI Coding Pipeline: A Portable Reference Architecture

Design Your Own Multi-AI Coding Pipeline: A Portable Reference Architecture

Comments
11 min read
You can't load-test an LLM agent with a dumb mock

You can't load-test an LLM agent with a dumb mock

Comments
3 min read
Stop writing a test-data builder for every class in .NET

Stop writing a test-data builder for every class in .NET

1
Comments
4 min read
Validating JSON-LD Beyond Syntax: Required Properties per schema.org Type

Validating JSON-LD Beyond Syntax: Required Properties per schema.org Type

Comments
6 min read
Reporting: Custom Reporters & Result Visibility (Playwright + TypeScript, Ch.25)

Reporting: Custom Reporters & Result Visibility (Playwright + TypeScript, Ch.25)

Comments
3 min read
testing the part of my chess app that downloads a 50mb binary

testing the part of my chess app that downloads a 50mb binary

Comments
4 min read
I Test Every AI Tool on Messy Data Before I Trust Any Demo

I Test Every AI Tool on Messy Data Before I Trust Any Demo

Comments
4 min read
Stability & Maintainability at Scale (Playwright + TypeScript, Ch.23)

Stability & Maintainability at Scale (Playwright + TypeScript, Ch.23)

Comments
3 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.