DEV Community

Testing

Find those bugs before your users do! 🐛

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
The Day I Became a Bug Hunter

Summer Bug Smash: Smash Stories 🐛🛹

The Day I Became a Bug Hunter

1
Comments
9 min read
How I cut my Chromatic bill 10x (works on any visual testing tool)

How I cut my Chromatic bill 10x (works on any visual testing tool)

5
Comments
4 min read
How Do You Build an Evaluation Harness for AI Agents?

How Do You Build an Evaluation Harness for AI Agents?

2
Comments 1
3 min read
Tenía 100% de cobertura y aun así se cayó producción

Tenía 100% de cobertura y aun así se cayó producción

Comments
6 min read
My Agent Said the Page Was Live. The Page Said 'We Are Closed.'

My Agent Said the Page Was Live. The Page Said 'We Are Closed.'

Comments
8 min read
How EvalPort's Grader System Works: 11 Types for LLM Evaluation

How EvalPort's Grader System Works: 11 Types for LLM Evaluation

Comments
2 min read
When a Failed Agent Step Looks Finished

When a Failed Agent Step Looks Finished

Comments
5 min read
My AI gate tests were green theater. The fix was to stub the wire — and nothing above it.

My AI gate tests were green theater. The fix was to stub the wire — and nothing above it.

Comments
5 min read
Fast Agent Quality Gates: Deterministic Rules Over LLM Judges

Fast Agent Quality Gates: Deterministic Rules Over LLM Judges

2
Comments 2
6 min read
A formula can be implemented perfectly and still give the wrong answer

A formula can be implemented perfectly and still give the wrong answer

Comments
5 min read
Three things Indian finance code gets wrong (with the numbers)

Three things Indian finance code gets wrong (with the numbers)

1
Comments
4 min read
The LLM was better at building a solver than playing the game

The LLM was better at building a solver than playing the game

Comments
5 min read
I built a BYOK AI agent that tests your app while you're building it, then turns the passing run into a real Playwright spec

I built a BYOK AI agent that tests your app while you're building it, then turns the passing run into a real Playwright spec

Comments
2 min read
Rewriting prose until the tests pass: everything passed, but the check that mattered never ran once

Rewriting prose until the tests pass: everything passed, but the check that mattered never ran once

Comments
6 min read
Three times this week a tool said it worked, and it had not

Three times this week a tool said it worked, and it had not

Comments
4 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.