Download for macOS
Skill

ab-test-analysis

@phuryn Updated 2026-07-03

Analyze A/B test results with statistical significance, sample size validation, confidence intervals, and ship/extend/stop recommendations. Use when evaluating experiment results, checking if a test reached significance, interpreting split test data, or deciding whether to ship a variant.

agent-skill-repositoryagent-skillsagentic-skillsclaude-code-marketplaceclaude-code-pluginsclaude-cowork-pluginproduct-management

Install

git clone https://github.com/phuryn/pm-skills /tmp/pm-skills && ln -s /tmp/pm-skills/pm-data-analytics/skills/ab-test-analysis ~/.claude/skills/ab-test-analysis

From README

A/B Test Analysis Evaluate A/B test results with statistical rigor and translate findings into clear product decisions. Context You are analyzing A/B test results for $ARGUMENTS. If the user provides data files (CSV, Excel, or analytics exports), read and analyze them directly. Generate Python scripts for statistical calculations when needed. Instructions Understand the experiment: What was the hypothesis? What was changed (the variant)? What is the primary metric? Any guardrail metrics? How long did the test run? What is the traffic split? Validate the test Sample size: Is the sample large enough for the expected effect size? Use the formula: n = (Z²α/2 × 2 × p × (1-p)) / MDE² Flag if the test is underpowered (<80% power) Duration: Did the test run for at least 1-2 full business cycles? Randomization: Any evidence of sample ratio mismatch (SRM)? Novelty/primacy effects: Was there enough time to wash out initial behavior changes?