Test prioritisation scoring
Use when you have a backlog of test ideas and need a defensible running order.
Fill in before running
Replace each placeholder with your own detail. The more specific you are, the less the model invents.
- {{BACKLOG}}
- {{TRAFFIC_BY_PAGE}}
- {{CAPACITY}}
- {{BUSINESS_PRIORITY}}
Getting a better result
- Feed in the raw backlog including the bad ideas - the scoring is what filters them.
- The "just ship it" list is often a quarter of the backlog and needs no test time.
- Re-score quarterly; evidence scores change as research lands.
Questions about this prompt
When do I score the backlog rather than pick the obvious next test?
When the backlog is longer than the quarter and the running order is being set by whoever argues hardest. It scores evidence, reach, expected effect, effort and learning value, and caps evidence at two for anything supported only by opinion or by a competitor doing it. That cap does most of the filtering.
What do I need in front of me before running it?
The raw backlog with whatever evidence each idea carries, traffic per page, team capacity and the current business priority. Feed in the weak ideas as well, since the scoring is what filters them and pruning first hides how thin the list really is. It will not invent traffic figures for pages you leave out.
What comes back, and which part is worth acting on?
A scored table sorted by total, the top three in run order with collisions flagged where two tests share a page, ideas the traffic cannot support, and ideas to ship without a test. That last list is often a quarter of the backlog and needs no test time, so read it before the rankings.
What is the mistake that costs me here?
Scoring once and treating the order as settled. Evidence scores move as research lands, so re-score quarterly. The other error is running two tests on the same page because both scored well: the prompt flags the collision, and ignoring it means neither result can be attributed cleanly to either change.