Skip to main content
Start projectENFR

Measurement and optimisation

Improvement as a standing process

Operloom sets up the research, prioritisation and testing practice that turns conversion work into a repeatable process rather than a series of opinions.

Redesign by opinionTests without powerNo hypothesis recordedWins that do not persist

01Business problems

Most testing programmes fail on the arithmetic

Underpowered tests read as wins, get shipped, and quietly deliver nothing. The discipline matters more than the tool.

  • Tests called before they reach significance
  • Traffic too low for the effect being sought
  • No recorded hypothesis, so nothing is learned
  • Results not segmented by device or source
  • Winners never re-checked after launch

02Research

Find the friction before changing anything

Research narrows the field so the testing budget is spent where the evidence points.

  1. 01Funnel and drop-off analysis
  2. 02Session review and form analytics
  3. 03Usability testing on the key journeys
  4. 04Sales and support call themes
  5. 05Accessibility and performance as conversion factors

03Method

A hypothesis, a power calculation, a decision rule

Every test states what it expects, how long it needs and what result would change the decision, before it starts.

  • Written hypothesis and expected mechanism
  • Sample size and duration calculated up front
  • Decision rule agreed before launch
  • Segmented analysis planned in advance
  • Result recorded whether it won or lost

04Where the sample is small

When testing is not the right instrument

Below a certain volume, split testing cannot resolve realistic effects. Saying so is more useful than running a test that cannot answer the question.

  1. 01Sequential and qualitative methods
  2. 02Fixing known defects rather than testing them
  3. 03Directional research with usability testing
  4. 04Improving measurement before optimising
  5. 05Being explicit about what cannot be proven

Common questions

Question: How much traffic is needed to test properly?
Operloom replies:

It depends on the baseline rate and the effect worth detecting, which is why every test starts with a power calculation. Where the numbers do not support a test, that is said plainly and a different method is used.

Related capability

Where this connects

Next step

Build a systemyour team can actually run

Start with the process as it works today, the platforms already in play and the outcome the organisation needs. We will tell you what we would change and in what order.