Cover the journeys you never had time to script.
Write the goal in plain language. The agents drive a real device and a real browser to reach it, and report every dead button, broken price and empty title they pass on the way. What passes becomes a replay you can schedule.
Opened the cart and tried the advertised code
Four things you get from one run
One goal, both surfaces
The same sentence of intent given to a real device and a real browser. Two runs, one vocabulary, and a report that says which surface each defect was on.
A run you did not write
The agents work out the route themselves. Nothing is pinned to a selector or a resource id, so a redesign does not red the suite the next morning.
The hygiene nobody scripts
Dead buttons, NaN prices, placeholder copy still live, stale copyright years, empty titles. The things a script would never think to check.
A replay when it passes
Promote a passing run and it re-runs exactly, with no model call at all. Classic automation reliability, none of the upkeep.
Four terms do most of the work. None of them are ours to redefine in a meeting later.
Goal
The outcome you want reached, written the way you would say it out loud. "Check out as a guest." No script, no SDK.
Test Atlas
The living, editable map of your app that every run merges into, and that every later run starts from.
Deterministic replay
A passing run promoted to a flow that repeats exactly the same way, with no model call in it at all.
Defect and friction
Two different tiers. A defect is the app getting something wrong. Friction is the app costing someone effort while being technically correct.
Goals, not selectors
You describe the outcome. Nothing in the run is pinned to a class name or a resource id, so a redesign does not break the suite the morning after it ships.
The same goal on both surfaces
A device run and a browser run of one sentence, reported in one vocabulary. A defect that only exists on one of them is the point, not an inconvenience.
Deterministic replay
Promote a run that passed and it replays exactly, with no model call at all. Classic automation reliability, none of the upkeep.
Test Atlas
Every run merges into one living map of your app. Edit it, run against it, and let each future run start from what the last one learned.



