GTMGO TO MARKETPLAYBOOKS
← All toolsEXPERIMENTATION · FREE TOOL

GTM experiment tracker

Set the sample, outcome thresholds, and cost guardrail before a GTM test. Record actuals later and export a decision brief.

How to use this tracker

Count only eligible exposures that could produce the primary outcome. Set a stop rate below the scale rate, and a minimum sample before interpreting results. The recommendation applies your own rules; it is not a causal or statistical test.

Inputs stay in this browser tab and are not sent to the site. Use fictional or approved information, never confidential employer or customer material.

01 · TEST DESIGN

Name the choice.

02 · PRECOMMIT

Set the rules.

Write these before reading the test result. Rates use qualified outcomes divided by eligible exposures.

03 · OBSERVED

Enter results.

Leave the precommitted rules intact. Count only outcomes that meet your definition above.

04 · Review the evidence

Three fictional results, three different calls

Each test uses 200 minimum eligible exposures, a 2% stop rate, a 5% scale rate, and an $80 cost-per-outcome limit. These are teaching values, not recommended benchmarks.

TestExposures / outcomesSpendRate / costRule check
Focused workflow guide240 / 14$5605.83% / $40Scale candidate
Broad message240 / 4$2401.67% / $60Stop candidate: weak qualified rate
Expensive channel240 / 14$1,4005.83% / $100Stop or redesign: cost limit breached

For the broad-message failure, inspect whether eligible people saw a clear, relevant offer before changing the threshold. For the cost failure, investigate acquisition cost and outcome quality; meeting the rate rule does not excuse a broken cost guardrail.

Keep the result interpretable

  1. Write the exposure unit. A delivered message, a unique eligible account, and a landing-page session are different denominators. Choose one before the test.
  2. Count the qualified outcome once. Deduplicate repeat actions and check the agreed qualification definition. Record missing and rejected outcomes separately.
  3. Include a comparable cost scope. Decide whether spend includes media, production, and delivery effort. Use the same scope for the threshold and actual result.
  4. Record contamination. If the offer, audience, or measurement changes, mark it in the review. Split the result or restart; do not quietly merge incompatible tests.
  5. Close the learning loop. Save the result, the alternative explanations, the next action, and the new review trigger. A failed test can still answer a useful question.

Download blank results worksheet ↓ Open the full experiment playbook →

What the tracker does not prove

This tool applies your decision rules. It does not calculate statistical significance, establish incrementality, adjust for channel selection bias, or recommend universal conversion thresholds. A minimum sample is a planning rule; it is not a statistical power calculation. Compare alternatives with a suitable design when a causal claim matters.