Drop in a few runs, get the picture
| # | Test | Flakiness | Runs (pass/fail) | Likely cause |
|---|
A test failing every run is broken, not flaky — FlakeRadar separates them so you fix the right thing. Always-pass tests aren't shown.
How the score works
Flip rate
How often a test switches pass↔fail between consecutive runs — the core instability signal. Pure alternation scores highest.
Balance & failure ratio
A 50/50 pass-fail split is flakier than a 1-in-10 blip; a 10/10 fail is broken, not flaky (scored separately).
Root-cause taxonomy
Error text is matched against an 8-category taxonomy — timing, order-dependency, data collision, environment, and more.
Quarantine & fix
Export a ready-to-paste quarantine list to green your pipeline today, then fix by category with the PRO RCA guidance.
Free to scan. $29 to fix like a pro.
- Unlimited in-browser analysis
- Flakiness scoring & ranking
- Broken-vs-flaky separation
- Basic cause hints + CSV export
- Full 8-category RCA engine (cause · why · fix)
- Quarantine-list export (Playwright/JUnit/pytest/Cypress)
- The flaky-test taxonomy reference
- JSON export + trend across sessions
- Lifetime license, this site
- Auto-ingest from GitHub Actions / Jenkins
- Hosted trend dashboard + history
- Weekly Slack / email digest
- Post-launch roadmap — join the list