# Analysis plan (written 2026-09-27, before any gold-standard label or outcome statistic was seen) The owner decided not to register on OSF. This file takes that role: it fixes the decision rules before the annotations and outcome analyses exist. `drafts/listicles/preregistration.md` remains the full specification of estimands and models. Any deviation is reported in the study's methodology as a deviation. ## Annotation Two independent model annotators (A: Claude Opus; B: Claude Sonnet) code all 299 gold items, blind to the detector. Disagreements are adjudicated by the analyst (Claude) reading the page text, with each decision logged in `annotations/listicles/adjudication.csv`. Reported as **model-coded validation, not human validation**. ## Choosing the parsing rule (dev half only) Candidates for numbered-heading pages: R-a (v1), R-b, R-c, and R-d-first (ItemList when present, else R-b). Publisher names: domain label only (v1) vs label + page aliases. 1. Primary criterion: balanced accuracy (mean of sensitivity and specificity) of **self-first** classification over all dev items in the numbered strata, weighted by stratum weights. 2. Tie-break (within 0.02): higher accuracy of the first entry among gold ranked lists. 3. Then prefer the simpler rule, in the order R-a, R-b, R-c, R-d-first, and label-only over aliases. The chosen rule is frozen and reported on the test half (sensitivity, specificity, PPV, with Wilson CIs, weighted). ## Coverage extension (sensitivity, not headline) For heading-only and no-structure pages: ItemList, then `