Fleet test coverage plane

No app can judge
its own tests

Every fleet app reports green about itself. This one scans them all from outside with the same yardstick, so lost coverage shows up as a regression, not a guess.

What it does

Six jobs no single app can do about itself

Each one answers a different part of "is this code actually proven?" — across the fleet, every night.

Scan

Nightly, each app's default branch runs in a throwaway worktree with a probe that records which routes, Livewire actions and methods every test reaches.

Judge

Every surface is marked Tested, Weak or Missing. Tested methods get their body mutated — a surviving mutant downgrades the verdict to Weak.

Ratchet

Each scan diffs against the last. Lost coverage and new untested surfaces are regressions — named, not averaged away.

Build

An AI agent writes tests for gaps, verifies them itself — reach, mutation, full suite — then opens one PR per app for a person to review.

Library

Every test of every app — read straight from git, with its revision history, outcome over 30 scans and how many surfaces it proves.

Proving Ground

Scripted personas walk the real UI in a headless browser. Failed steps, server errors and clicks with no loading feedback become findings.

End to end

One customer, start to finish

Every stage of a customer's life — from first enquiry to cancellation — walked across every app it touches in a stand-in fleet.

Enquiry

Built

Payment

Built

Subscription

Built

Provisioning

Built

First sign-in

Built

First briefing

Pending

Upgrade

Built

Cancellation

Built