跳到主要内容
知仓学习社ZHICANG

aer-preregistration

Use when the project collects primary data or runs a field, lab, or survey experiment, before the intervention begins — write the pre-analysis plan,…

不碰外部(只输出文字)无严重或高危命中brycewang-stanford/Auto-Empirical-Research-Skills

它会碰到什么

扫了多少3 个文本文件,17 KB
它会碰到什么不碰外部(只输出文字)
命中总数0 处
命中统计严重 0 · 高 0 · 中 0 · 低 0

这一栏是扫描器报的事实,不是结论。命中多不等于有毒(安全工具、规则库、示例脚本本来就会包含危险写法),命中少也不等于干净。它和你手上的凭据、文件、网络有什么关系,需要你自己看。

技能内容

AER Pre-Registration

Overview

For experimental and prospective-data AER-track work, credibility is bought

before the data exist. This skill writes the pre-analysis plan (PAP), sizes

the sample from a power calculation, and registers the study. Its job is to make

the eventual results un-p-hackable: a referee who sees a public PAP timestamp

predating data collection cannot accuse you of specification search.

Registration is not optional at AEA journals for RCTs. If the design is

observational, skip to aer-robustness; there is nothing to pre-register.

When to Use

  • The paper runs a randomized controlled trial (field, lab, or survey experiment)
  • The project collects primary data whose analysis should be pre-committed
  • A referee or editor asks for the trial registration number or the PAP
  • Reviewers suspect the reported specification was chosen after seeing outcomes
  • You must justify the sample size, or explain why an effect was "not detected"

Register-or-Not Decision

Is treatment assigned by the researcher (randomization)?
├── Yes → register with the AEA RCT Registry BEFORE the intervention; write a PAP
├── No, but you collect new primary data → post a PAP for the pre-committed analysis
└── No, secondary/observational data already realized → do NOT pre-register
      (a PAP written after the outcomes exist is theater); go to aer-robustness

Pre-Analysis Plan Contents

Pre-specify, in order, and timestamp before unblinding:

  1. Hypotheses — each stated as a signed, testable prediction, not a topic:

> H1: the cash transfer raises household consumption at endline by at

> least 0.15 SD, estimated by OLS of log consumption on treatment with

> strata fixed effects, clustered at the village level.

  1. Primary outcomes — a short list (ideally 1-3). Everything else is

secondary or exploratory and labeled as such.

  1. Estimating equation — the exact regression, fixed effects, covariates,

and the level at which standard errors cluster.

  1. Sample and inclusion rules — who is in the analysis and how attrition is

handled (pre-commit to bounds; see examples/lee-bounds-demo/).

  1. Multiple-testing correction — the family and the method (e.g. Romano-Wolf,

romano_wolf_2005), pre-specified, not chosen after the p-values land.

  1. Heterogeneity — the subgroups you will test, fixed in advance; all others

are exploratory.

  1. Power — the MDE and the assumptions behind it (below).

Keep the PAP moderate in scope (Olken's advice): pre-specify the primary

analysis tightly, leave genuine discovery clearly flagged as exploratory. An

over-long PAP that pre-registers forty outcomes protects nothing.

Power and the Minimum Detectable Effect

Size the sample from the MDE, not the other way around. For a two-arm trial with

equal allocation, size sigma, share p, total N:

MDE = (z_power + z_{1-alpha/2}) * sigma * sqrt(1 / (p (1 - p) N))
  • Target 80% power at a 5% two-sided level: $z_{0.80} = 0.84$ and

$z_{0.975} = 1.96$, so z_power + z_alpha = 2.80.

  • State the MDE in the units the audience cares about, and benchmark it against

the smallest effect that would be economically interesting. If the MDE exceeds

that, the study is underpowered — do not run it as designed.

  • Correct the variance for clustered assignment (design effect

$1 + (m-1)\rho$ — e.g. $m = 30$ per cluster and $\rho = 0.05$ inflates the

needed sample by a factor of $2.45$), for a baseline covariate ($R^2$ gain),

and for expected attrition. A power number that ignores the ICC is fiction.

  • Underpowered designs do not just miss: when they reach significance they

exaggerate the effect (Type-M). See examples/power-mde-demo/ for the

MDE-attains-target-power check and the winner's-curse simulation.

Cite mckenzie_2012 (more rounds beat larger cross-sections when outcomes are

noisy) and duflo_glennerster_kremer_2007 (the design toolkit). Keys in

[../../references.bib](../../references.bib); defaults in

[../../docs/methods-reference.md](../../docs/methods-reference.md).

Reporting Against the Plan

  • Report the pre-specified primary result first, exactly as written, even if

it is null. A pre-registered null is a publishable finding, not a failure.

  • Mark every deviation from the PAP explicitly, with the reason, in a table.
  • Separate confirmatory from exploratory results in the manuscript; never

promote an exploratory subgroup to the headline.

  • File the disclosure and IRB/ethics approvals; AEA journals require both.

Red Flags for Referees

  • No registration number on an RCT, or a timestamp after data collection began
  • A power calculation that omits clustering, attrition, or the ICC
  • Twenty "primary" outcomes with no multiplicity correction
  • The headline result is an unregistered subgroup interaction
  • "Not statistically significant" reported as "no effect" with no MDE stated
  • Deviations from the PAP that are silent rather than disclosed

Pre-Registration Gate

Do not advance to data collection until all are true:

  • [ ] Primary outcomes and the estimating equation are pre-specified in writing
  • [ ] The sample size is justified by an MDE with clustering and attrition built in
  • [ ] The multiple-testing family and correction are fixed in advance
  • [ ] The study is registered (AEA RCT Registry) with a pre-intervention timestamp
  • [ ] Exploratory analyses are labeled as such, not disguised as confirmatory

Repository Resources

Bundled with the installed skill, no repository checkout needed --- read it

before the repo resources below:

  • references/pap-template.md --- PAP outline, power/MDE reporting template, registry field checklist

When working from the repo or plugin bundle, load only the relevant resource:

  • Power/MDE worked simulation with the Type-M winner's-curse check: examples/power-mde-demo/
  • Attrition bounds to pre-commit in the PAP: examples/lee-bounds-demo/
  • Estimator defaults, diagnostics, and BibTeX keys: docs/methods-reference.md
  • Multiple-testing correction methods: skills/aer-robustness/SKILL.md
  • Verified references (mckenzie_2012, duflo_glennerster_kremer_2007): references.bib

Fix the MDE and the primary-outcome list before drafting; both feed the

aer-identification estimator choice and the aer-consistency audit.

Handoff

DESIGN: <RCT | primary-data collection | observational (no PAP)>
PRIMARY OUTCOMES: <list, 1-3>
MDE / POWER: <MDE in outcome units; power; assumptions incl. ICC and attrition>
MULTIPLICITY: <family + correction method>
REGISTRATION: <AEA RCT ID + timestamp, or "n/a">
NEXT SKILL: aer-identification (confirm estimator) then aer-robustness

Anti-Patterns

  • Writing the PAP after the endline data are in hand — the timestamp is the point
  • Pre-registering so many outcomes that "confirmatory" loses all meaning
  • Powering the study for the effect you hope for instead of the smallest one

worth detecting

  • Treating a registered null as a failed experiment rather than a clean result
  • Reporting an exploratory subgroup as if it had been pre-specified

想直接用这个技能?

本站把开放许可(MIT / Apache 等)的技能按仓库打包整理到网盘,点一下转存到你自己的网盘,不用一个个从 GitHub 拉。许可未声明的技能只给原始仓库链接,不打包。