Results dashboard
Stored runs per project, over time: whether the suite is getting better at catching bugs or just running more lines, which kinds of bugs survive, and what each fix run changed. Every value is read from the backend as measured.
Line coverage…No runs yet
Mutation score…No runs yet
Coverage − mutation…No runs yet
Quality gate…No runs yet
Coverage vs. mutation score per commit
Is the suite getting better at catching bugs, or just running more lines?
/trendLoading…
Survival by mutation operator
Which kinds of bugs slip through?
/operatorsLoading…
Before / after each fix run
Did the AI-written tests actually move the score, and by how much?
/fix-effectLoading…
File risk over time
Where is risk concentrating?
/risk-heatmapLoading…
Untested Endpoints
Which API routes lack test coverage?
/endpointsLoading…
Persistent survivors
Which weak spots have never been caught?
/survivorsLoading…
Flaky tests
Which tests are untrustworthy signals?
/flakyLoading…