#methodological rigourPooling robustness remains incompletely tested
The four projects differ in replication criteria, power, and forecasting conditions, and estimated heterogeneity is substantial. The headline pooled estimates are not accompanied by a formal leave-one-study-out sensitivity analysis.
↳ Methods §2.1; Results §3.1
#positioningForecaster selection limits broader inference
Participants were recruited through blogs, mailing lists, and Twitter, but the implications of this self-selected group for claims about the scientific community are not fully traced in the Discussion.
↳ Methods §2.2; Discussion
#reportingMinor inconsistencies require correction
Section 3.3 reports the variance-weighted mean as M = 58 and SD = 17 rather than proportions, while RPP accuracy appears as 70% in §2.2.1 and 71% in §3.2.
↳ Methods §2.2.1; Results §§3.2–3.3