What you get
• Sample size per arm and estimated runtime, with the assumptions made explicit.
• Minimum detectable effect and power analysis to assess whether the planned test can detect an improvement worth acting on.
• Guardrail metric recommendations so a conversion lift does not hide a costly tradeoff.
• A pre-registration summary covering the hypothesis, metrics and stopping rule.
• A written recommendation: proceed, revise the design, or reconsider the test given the available traffic.