Results

Results dashboard

Stored runs per project, over time: whether the suite is getting better at catching bugs or just running more lines, which kinds of bugs survive, and what each fix run changed. Every value is read from the backend as measured.

Line coverage…No runs yet
Mutation score…No runs yet
Coverage − mutation…No runs yet
Quality gate…No runs yet
1

Coverage vs. mutation score per commit

Is the suite getting better at catching bugs, or just running more lines?

Two lines, shaded gap/trend
Loading…
2

Survival by mutation operator

Which kinds of bugs slip through?

Sorted horizontal bars/operators
Loading…
3

Before / after each fix run

Did the AI-written tests actually move the score, and by how much?

Slope chart/fix-effect
Loading…
4

File risk over time

Where is risk concentrating?

Heatmap, file × run/risk-heatmap
Loading…
5

Untested Endpoints

Which API routes lack test coverage?

List/endpoints
Loading…
6

Persistent survivors

Which weak spots have never been caught?

Table/survivors
Loading…
7

Flaky tests

Which tests are untrustworthy signals?

Table/flaky
Loading…