跳到主要内容
知仓学习社ZHICANG

ab-test-setup

>

读文件无严重或高危命中borghei/Claude-Skills

它会碰到什么

扫了多少5 个文本文件,45 KB
它会碰到什么读文件
命中总数2 处
命中统计严重 0 · 高 0 · 中 2 · 低 0

这一栏是扫描器报的事实,不是结论。命中多不等于有毒(安全工具、规则库、示例脚本本来就会包含危险写法),命中少也不等于干净。它和你手上的凭据、文件、网络有什么关系,需要你自己看。

技能内容

A/B Test Setup Skill

Overview

Production-ready A/B testing toolkit for calculating sample sizes, designing rigorous test plans, and analyzing results with statistical significance testing. Designed for growth teams, product managers, and marketers who need to make data-driven decisions from controlled experiments.

Clarify First

Before designing the test, confirm these inputs. If any is unknown or vague, ASK — do not assume:

  • [ ] Hypothesis + primary metric — what change you expect and the single metric that judges it (drives test plan + analysis)
  • [ ] Baseline conversion rate — the current rate the metric sits at today (drives sample size calculation)
  • [ ] Minimum detectable effect (MDE) — smallest lift worth detecting (drives required samples + duration)
  • [ ] Daily traffic available — eligible visitors per day per variant (determines how long the test must run)

Stop rule: ask only the 2-3 that most change the output. If the user says "just draft it," proceed and list your assumptions at the top of the artifact.

Quick Start

# Calculate required sample sizes for a test
python scripts/sample_size_calculator.py --baseline 0.05 --mde 0.10 --power 0.80

# Design a complete A/B test plan
python scripts/test_designer.py test_config.json

# Analyze A/B test results
python scripts/results_analyzer.py results.json

Tools Overview

| Tool | Purpose | Input | Output |

|------|---------|-------|--------|

| sample_size_calculator.py | Sample size calculation | Baseline rate, MDE, power | Required samples + duration |

| test_designer.py | Test plan design | JSON test config | Complete test plan document |

| results_analyzer.py | Results analysis | JSON with test results | Statistical analysis + recommendation |

Workflows

Workflow 1: New A/B Test Setup

  1. Define hypothesis and success metric
  2. Run sample_size_calculator.py with baseline conversion and minimum detectable effect
  3. Create test configuration JSON (see Common Patterns)
  4. Run test_designer.py to generate complete test plan
  5. Share plan with stakeholders for alignment before launch

Workflow 2: Test Results Analysis

  1. Collect test results into JSON format
  2. Run results_analyzer.py to get statistical significance
  3. Review confidence interval, p-value, and effect size
  4. Check for segment-level effects if overall result is inconclusive
  5. Make ship/no-ship decision based on analysis

Workflow 3: Experimentation Program Review

  1. Compile results from multiple past tests
  2. Run results_analyzer.py --batch on all results
  3. Review win rate, average effect size, and velocity
  4. Identify patterns in winning vs losing tests
  5. Optimize test pipeline based on learnings

Reference Documentation

See references/ab-testing-guide.md for comprehensive methodology covering:

  • Statistical foundations (z-tests, confidence intervals)
  • Sample size theory and trade-offs
  • Common experimentation pitfalls
  • Multi-variant and sequential testing
  • Bayesian vs frequentist approaches

Common Patterns

Pattern: Test Configuration JSON

{
  "test_name": "Homepage CTA Button Color",
  "hypothesis": "Changing the CTA button from blue to green will increase click-through rate",
  "metric_primary": "cta_click_rate",
  "metric_secondary": ["signup_rate", "bounce_rate"],
  "baseline_rate": 0.045,
  "minimum_detectable_effect": 0.10,
  "significance_level": 0.05,
  "power": 0.80,
  "variants": [
    {"name": "control", "description": "Current blue CTA button"},
    {"name": "treatment", "description": "Green CTA button"}
  ],
  "daily_traffic": 5000,
  "allocation": {"control": 0.50, "treatment": 0.50}
}

Pattern: Test Results JSON

{
  "test_name": "Homepage CTA Button Color",
  "variants": {
    "control": {"visitors": 12500, "conversions": 563},
    "treatment": {"visitors": 12500, "conversions": 625}
  },
  "metric": "cta_click_rate",
  "significance_level": 0.05
}

Quick Reference: Common Effect Sizes

| Context | Small Effect | Medium Effect | Large Effect |

|---------|-------------|---------------|--------------|

| Conversion Rate | 2-5% relative | 5-15% relative | > 15% relative |

| Revenue per User | 1-3% | 3-8% | > 8% |

| Engagement Rate | 3-5% | 5-10% | > 10% |

想直接用这个技能?

本站把开放许可(MIT / Apache 等)的技能按仓库打包整理到网盘,点一下转存到你自己的网盘,不用一个个从 GitHub 拉。许可未声明的技能只给原始仓库链接,不打包。

同名技能的其他版本

有 2 个不同仓库或目录里都有叫 ab-test-setup 的技能。它们内容并不相同,别混用: