experiment-readout

Installation
SKILL.md

Experiment Readout Skill

A test result is only a decision if the statistics are sound — and "variant looks higher" is not a result. This skill computes the lift, the p-value, and a confidence interval from the raw counts, checks the guardrails, and writes an honest readout with a clear ship/no-ship call — flagging the traps (peeking, underpowered, novelty, a significant but tiny effect) that make teams ship noise.

Required Inputs

Ask for these only if they aren't already provided:

  • The metric & data — for a conversion test: users and conversions per variant (control vs. treatment). For a continuous metric: mean, SD, and n per variant.
  • The hypothesis — what you expected and the minimum effect that matters.
  • Guardrail metrics — what shouldn't get worse (revenue, latency, retention).
  • Test setup — planned sample size/duration, and whether it ran to plan (for the peeking check).

Output Format

Experiment Readout: [test name]

Installs
4
GitHub Stars
1.2K
First Seen
Jun 28, 2026
experiment-readout — mohitagw15856/pm-claude-skills