Experiment Desk — results in, a decision-ready experiment readout out
Paste a finished A/B test and get the readout the decision meeting needs. The browser does the whole statistical pass for free: the two-proportion z-test (pooled standard error for the p-value, unpooled for the interval) or Welch's test, confidence intervals on the absolute and relative lift, the power the design achieved, the sample size its declared MDE required, the smallest effect the collected exposure could have detected, the sample-ratio-mismatch chi-square, the whole-business-cycle duration check, the peeking penalty with the Pocock boundary, a Holm correction across the guardrail family, and a decision reached by rule — ship, investigate, extend, stop, revert or invalid. When the result is inconclusive it solves for what would settle it: the exposure per arm and the extra days at your own measured traffic rate. The AI pass adds only judgement — whether the hypothesis was testable, what a flat metric means for the product, the validity threats the arithmetic cannot see, and the next experiment — and every reading it gives is counted exactly once against the measured metrics and checked against the measured significance. A derived work of the @phuryn/ab-test-analysis skill.
Details
gpt-terra Every public app is built from a security-scanned skill and must pass a clean scan — skill and frontend — before it can be listed. Have a skill of your own? Turn it into an app — or read the step-by-step walkthrough.