After the two 2026 preprints went out, a post-hoc look at the Part 1 response counts turned up a systematic pattern: clips presented later in a wave's fixed order picked up meaningfully fewer responses than clips at the start of the same wave. Across the three collection waves, the within-wave rank correlation between clip position and response count was consistently strong and negative (Spearman ρ = −0.78 to −0.92; p ≤ 0.005 in each wave; ρ = −0.43 pooled, p = 0.012).
Two things follow from that. First, the Part 1 dataset used in the two preprints has a real dropout confound: later-position clips carry less statistical weight than their earlier neighbours, and any comparison between clips on the tail of a wave and clips on the head of a wave is standing on uneven ground. Second, this is a fixable problem, but only for future data — the fair move is to lock the Part 1 dataset exactly as the papers used it and to change the mechanism for the next round.
What we changed: The Part 1 dataset is now version 1.0 as of 14 Aug 2026, published as a Zenodo deposit (doi.org/10.5281/zenodo.22265192). Phase 1 · Part 2, opening in September 2026, keeps the same 33 clips and the same response instrument, but (a) randomizes clip order per session within each wave and (b) gates parental-consent verification at session entry rather than as a post-hoc exclusion. A subsequent definitive Phase 1 paper will draw on the Part 2 corpus.