Skip to content

Topic

Flaky And Seed-Dependent Tests

Tests whose pass/fail result varies across runs due to randomness, timing, or small sample sizes, undermining confidence in CI results.

Current stories

build1 publisher

Split the suite so flaky tests cannot gate an agent's patch

A dev.to write-up sorts every test path into four classes and hashes the scoring ones on main, so a patch that retouches a fixture fails with a different exit code than a patch that breaks a property.

Publishers:dev.to

Reality

Evidence32
Adoption
Insufficient
Hype gap+18
Incentives18
Confidence45
build1 publisher

A 50-run repro loop finds a one-in-100 flake about two times in five

A dev.to engineer's runbook treats flaky microservice tests as incidents to be measured and isolated before anyone reaches for a fix. Its fifty-run reproduction loop is what decides which flakes a team can diagnose at all.

Publishers:dev.to

Reality

Evidence38
Adoption
Insufficient
Hype gap+15
Incentives18
Confidence45