Shiyao Liu, Junni L. Zhang
arXiv 27 Aug 2026 · Statistics — Methodology
arXiv:2608.26606 · PDF · Extracted main text
Recent work encourages political scientists to move from post-only toward within-subject designs for improved precision from repeated measurements. We formalize a potential-outcomes framework for two-period within-subject designs that allows for unequal allocation and heterogeneous treatment and carryover effects. We characterize the pooled estimator and evaluate the carryover test used to justify pooling. We find: first, pooling identifies the average treatment effect only when the gap in the average carryover effects is zero across the two treatment sequences. The unit-clustered standard error for the pooled estimator is identical to its design-based counterpart. Second, under mild conditions, the carryover test has strictly less power than the average-treatment-effect test with post-only data. The resulting two-step procedure, which pools only after a nonrejected test, produces confidence intervals that typically undercover. When the gap is zero, undercoverage occurs if and only if pooling is more efficient than post-only analysis, precisely when the within-subject design is worthwhile. When the gap is nonzero, undercoverage is typical unless the gap or sample size is large. Third, we derive a sensitivity analysis and find published conclusions robust to plausible carryover gaps. We therefore endorse within-subject designs but recommend justifying a zero carryover gap substantively and reporting sensitivity to departures.
appendix boundary found by none_found · 100% of the source is main text. Read the extracted text to check this.
The works this paper leans on most, across its whole bibliography — not restricted to papers in our corpus. Ranked by composite intensity, which combines how often a work is mentioned, how many sections mention it, and how much of that falls in the main text rather than the appendix.
| Reference | Intensity | Mentions | Sections | Main text | |
|---|---|---|---|---|---|
| 1 | Clifford, Scott, Sheagley, Geoffrey, Piston, Spencer (2021) Increasing precision without altering treatment effects: Repeated measures designs in survey experiments | 1.000 | 27 | 7 | 100% |
| 2 | Jordan, Diana, Ollerenshaw, Trent, Trexler, Andrew (2026) New Evidence and Design Considerations for Repeated Measure Experiments in Survey Research | 1.000 | 16 | 5 | 100% |
| 3 | Freeman, PR (1989) The performance of the two-stage analysis of two-treatment, two-period crossover trials | 1.000 | 7 | 4 | 100% |
| 4 | Carnahan, Dustin, Bergan, Daniel E (2022) Correcting the misinformed: the effectiveness of fact-checking messages in changing false beliefs | 0.644 | 2 | 2 | 100% |
| 5 | Clifford, Scott, Sheagley, Geoffrey, Piston, Spencer (2021) Replication Data for: Increasing Precision without Altering Treatment Effects: Repeated Measures Designs in Survey Experiments | 0.644 | 2 | 2 | 100% |
| 6 | Halling, Aske (2024) Frontline employees' responses to citizens' communication of administrative burdens | 0.644 | 2 | 2 | 100% |
| 7 | Jordan, Diana, Ollerenshaw, Trent, Trexler, Andrew (2026) Replication Data for: New Evidence and Design Considerations for Repeated Measure Experiments in Survey Research | 0.644 | 2 | 2 | 100% |
| 8 | Ozer, Adam L., Wright, Jamie M (2022) Partisan news versus party cues: The effect of cross-cutting party and partisan network cues on polarization and persuasion | 0.644 | 2 | 2 | 100% |
| 9 | Tappin, Ben M (2023) Estimating the between-issue variation in party elite cue effects | 0.644 | 2 | 2 | 100% |
| 10 | Van Trappen, Sigrid (2023) Biased expectations? An experimental test of which party selectors are more likely to stereotype ethnic minority aspirants as le… | 0.644 | 2 | 2 | 100% |
Showing the top 10 of 44 scored citations.