Blanket licensing reduces self-selection bias
Articles were opened through publisher-library licensing rather than author choice, directly addressing a prominent confound in prior OA citation studies.
↳ Methodology, opening paragraphs
Loading evaluation data…
Many studies show that open access (OA) articles-articles from scholarly journals made freely available to readers without requiring subscription fees-are downloaded, and presumably read, more often than closed access/subscription-only articles. Assertions that OA articles are also cited more often generate more controversy. Confounding factors (authors may self-select only the best articles to make OA; absence of an appropriate control group of non-OA articles with which to compare citation figures; conflation of pre-publication vs. published/publisher versions of articles, etc.) make demonstrating a real citation difference difficult. This study addresses those factors and shows that an open access citation advantage as high as 19% exists, even when articles are embargoed during some or all of their prime citation years. Not surprisingly, better (defined as above median) articles gain more when made OA.
the results lend support to the existence of a real, measurable, open access citation advantage with a lower bound of approximately 20%
heterogeneous comparator estimates include smaller positive, null, and negative results
an open access citation advantage as high as 19% exists, even when articles are embargoed during some or all of their prime citation years
the favourable median estimate conflicts with negative and smaller alternative estimators
Not surprisingly, better (defined as above median) articles gain more when made OA.
direction is generally consistent, though closest-equivalent and causal interpretations remain uncertain
That large advantage shrinks when articles are treated individually and compared to close equivalents, but it doesn't disappear.
article-level closest-equivalent analysis reports a smaller positive estimate
Derived from the full evaluation — not a separate score.
Strengths
Articles were opened through publisher-library licensing rather than author choice, directly addressing a prominent confound in prior OA citation studies.
↳ Methodology, opening paragraphs
The Analysis presents mean, median, equivalent, closest-equivalent, and newer-article results, including negative and nonsignificant estimates rather than only favourable comparisons.
↳ Analysis, Overall and refinement results
The Discussion identifies institutional concentration, disciplinary imbalance, article age, embargo timing, and annual citation granularity, including the likely direction of timing-related measurement error.
↳ Discussion, limitations paragraphs
Limitations
For the overall sample, the mean estimate is −7.6%, the median estimate is 18.5%, and the equivalent-article estimate is 3.6%. This sign and magnitude variation prevents a stable numerical interpretation.
↳ Analysis, Overall n=3,850 percentage results
The conclusion asserts an approximately 20% lower bound despite a 10.7% closest-equivalent estimate and several null or negative comparisons. The wording materially exceeds the reported evidentiary range.
↳ Analysis; Conclusion, final paragraph
All opened articles came from one institutional repository, and 92% were in physical science, health science, and engineering. The paper itself identifies multi-institutional evidence as preferable.
↳ Methodology; Discussion, limitations
The blanket-licensing setting and use of same-issue comparisons provide a meaningful design improvement over author-selected repository studies. The Analysis is unusually transparent about conflicting estimators, including negative and nonsignificant results, but that variation also places the methodological execution in the 2.5–3.4 range. The strongest downward pressure comes from the conclusion’s approximately 20% lower bound, which is not supported consistently by the mean, equivalent, closest-equivalent, or newer-article analyses. The study is therefore useful cumulative evidence for a possible modest advantage, not a stable causal estimate suitable for direct policy use.
Nabu’s assessment, alongside the field’s view.
Are you an author of this paper?
Sound3.3
Confidence mediumThe blanket-licensing design meaningfully extends prior OA citation research by reducing author self-selection and comparing final published versions over long observation windows. Its contribution remains incremental because the principal estimates vary substantially by comparator.
None of the OA articles were self-selected
Same-issue controls, pre-opening citation histories, and explicit temporal ordering provide a credible observational comparison. Residual nonexchangeability, zero-inflated outcomes, multiple analyses, and comparator sensitivity introduce meaningful uncertainty about causal and numerical conclusions.
It is an imperfect proxy, of course
The progression from aggregate to article-level, closest-equivalent, and newer-article analyses is clearly presented, including null and negative results. The conclusion’s approximately 20% lower bound materially exceeds what the heterogeneous estimates establish.
a lower bound of approximately 20%
The introduction engages established OACA findings and methodological criticisms, while the discussion links several limitations to interpretation and future research. The literature base is relatively compact, and the broad concluding estimate is insufficiently narrowed by the single-institution and disciplinary constraints.
a multi-institution sample would be ideal
Caveats4 of 4 checks
The reported analyses are transparent but yield materially inconsistent estimates, and some statistical reporting warrants caution. These issues affect precision and interpretation without demonstrating that the central qualitative result reverses entirely.
The study uses published bibliometric data and therefore does not require human-participant ethics approval. It claims anonymized data availability through a named repository and DOI, with no internally suspicious conduct issue identified.
Flags: 1 declared / 5 total
9 of 9 checkable references verified
12 references in manuscript 3 are books, websites or datasets — counted, but not index-checkable citation details diverge from the cited records for 1 reference
No retraction notice found in Retraction Watch.
Sources: Retraction Watch ✓
Where this paper’s evidence sits on the path from initial observation to real-world use.
The study tests citation behavior using operational repository and journal data rather than a laboratory setting. Its observational design and estimator-dependent results require triangulation before they can guide access or embargo policy.
these somewhat equivocal results
AI-generated, human-governed. Something look off? Contact us to request a review.