Cargando…
A descriptive analysis of the data availability statements accompanying medRxiv preprints and a comparison with their published counterparts
OBJECTIVE: To determine whether medRxiv data availability statements describe open or closed data—that is, whether the data used in the study is openly available without restriction—and to examine if this changes on publication based on journal data-sharing policy. Additionally, to examine whether d...
Autores principales: | , |
---|---|
Formato: | Online Artículo Texto |
Lenguaje: | English |
Publicado: |
Public Library of Science
2021
|
Materias: | |
Acceso en línea: | https://www.ncbi.nlm.nih.gov/pmc/articles/PMC8118451/ https://www.ncbi.nlm.nih.gov/pubmed/33983972 http://dx.doi.org/10.1371/journal.pone.0250887 |
_version_ | 1783691753038020608 |
---|---|
author | McGuinness, Luke A. Sheppard, Athena L. |
author_facet | McGuinness, Luke A. Sheppard, Athena L. |
author_sort | McGuinness, Luke A. |
collection | PubMed |
description | OBJECTIVE: To determine whether medRxiv data availability statements describe open or closed data—that is, whether the data used in the study is openly available without restriction—and to examine if this changes on publication based on journal data-sharing policy. Additionally, to examine whether data availability statements are sufficient to capture code availability declarations. DESIGN: Observational study, following a pre-registered protocol, of preprints posted on the medRxiv repository between 25th June 2019 and 1st May 2020 and their published counterparts. MAIN OUTCOME MEASURES: Distribution of preprinted data availability statements across nine categories, determined by a prespecified classification system. Change in the percentage of data availability statements describing open data between the preprinted and published versions of the same record, stratified by journal sharing policy. Number of code availability declarations reported in the full-text preprint which were not captured in the corresponding data availability statement. RESULTS: 3938 medRxiv preprints with an applicable data availability statement were included in our sample, of which 911 (23.1%) were categorized as describing open data. 379 (9.6%) preprints were subsequently published, and of these published articles, only 155 contained an applicable data availability statement. Similar to the preprint stage, a minority (59 (38.1%)) of these published data availability statements described open data. Of the 151 records eligible for the comparison between preprinted and published stages, 57 (37.7%) were published in journals which mandated open data sharing. Data availability statements more frequently described open data on publication when the journal mandated data sharing (open at preprint: 33.3%, open at publication: 61.4%) compared to when the journal did not mandate data sharing (open at preprint: 20.2%, open at publication: 22.3%). CONCLUSION: Requiring that authors submit a data availability statement is a good first step, but is insufficient to ensure data availability. Strict editorial policies that mandate data sharing (where appropriate) as a condition of publication appear to be effective in making research data available. We would strongly encourage all journal editors to examine whether their data availability policies are sufficiently stringent and consistently enforced. |
format | Online Article Text |
id | pubmed-8118451 |
institution | National Center for Biotechnology Information |
language | English |
publishDate | 2021 |
publisher | Public Library of Science |
record_format | MEDLINE/PubMed |
spelling | pubmed-81184512021-05-24 A descriptive analysis of the data availability statements accompanying medRxiv preprints and a comparison with their published counterparts McGuinness, Luke A. Sheppard, Athena L. PLoS One Research Article OBJECTIVE: To determine whether medRxiv data availability statements describe open or closed data—that is, whether the data used in the study is openly available without restriction—and to examine if this changes on publication based on journal data-sharing policy. Additionally, to examine whether data availability statements are sufficient to capture code availability declarations. DESIGN: Observational study, following a pre-registered protocol, of preprints posted on the medRxiv repository between 25th June 2019 and 1st May 2020 and their published counterparts. MAIN OUTCOME MEASURES: Distribution of preprinted data availability statements across nine categories, determined by a prespecified classification system. Change in the percentage of data availability statements describing open data between the preprinted and published versions of the same record, stratified by journal sharing policy. Number of code availability declarations reported in the full-text preprint which were not captured in the corresponding data availability statement. RESULTS: 3938 medRxiv preprints with an applicable data availability statement were included in our sample, of which 911 (23.1%) were categorized as describing open data. 379 (9.6%) preprints were subsequently published, and of these published articles, only 155 contained an applicable data availability statement. Similar to the preprint stage, a minority (59 (38.1%)) of these published data availability statements described open data. Of the 151 records eligible for the comparison between preprinted and published stages, 57 (37.7%) were published in journals which mandated open data sharing. Data availability statements more frequently described open data on publication when the journal mandated data sharing (open at preprint: 33.3%, open at publication: 61.4%) compared to when the journal did not mandate data sharing (open at preprint: 20.2%, open at publication: 22.3%). CONCLUSION: Requiring that authors submit a data availability statement is a good first step, but is insufficient to ensure data availability. Strict editorial policies that mandate data sharing (where appropriate) as a condition of publication appear to be effective in making research data available. We would strongly encourage all journal editors to examine whether their data availability policies are sufficiently stringent and consistently enforced. Public Library of Science 2021-05-13 /pmc/articles/PMC8118451/ /pubmed/33983972 http://dx.doi.org/10.1371/journal.pone.0250887 Text en © 2021 McGuinness, Sheppard https://creativecommons.org/licenses/by/4.0/This is an open access article distributed under the terms of the Creative Commons Attribution License (https://creativecommons.org/licenses/by/4.0/) , which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited. |
spellingShingle | Research Article McGuinness, Luke A. Sheppard, Athena L. A descriptive analysis of the data availability statements accompanying medRxiv preprints and a comparison with their published counterparts |
title | A descriptive analysis of the data availability statements accompanying medRxiv preprints and a comparison with their published counterparts |
title_full | A descriptive analysis of the data availability statements accompanying medRxiv preprints and a comparison with their published counterparts |
title_fullStr | A descriptive analysis of the data availability statements accompanying medRxiv preprints and a comparison with their published counterparts |
title_full_unstemmed | A descriptive analysis of the data availability statements accompanying medRxiv preprints and a comparison with their published counterparts |
title_short | A descriptive analysis of the data availability statements accompanying medRxiv preprints and a comparison with their published counterparts |
title_sort | descriptive analysis of the data availability statements accompanying medrxiv preprints and a comparison with their published counterparts |
topic | Research Article |
url | https://www.ncbi.nlm.nih.gov/pmc/articles/PMC8118451/ https://www.ncbi.nlm.nih.gov/pubmed/33983972 http://dx.doi.org/10.1371/journal.pone.0250887 |
work_keys_str_mv | AT mcguinnesslukea adescriptiveanalysisofthedataavailabilitystatementsaccompanyingmedrxivpreprintsandacomparisonwiththeirpublishedcounterparts AT sheppardathenal adescriptiveanalysisofthedataavailabilitystatementsaccompanyingmedrxivpreprintsandacomparisonwiththeirpublishedcounterparts AT mcguinnesslukea descriptiveanalysisofthedataavailabilitystatementsaccompanyingmedrxivpreprintsandacomparisonwiththeirpublishedcounterparts AT sheppardathenal descriptiveanalysisofthedataavailabilitystatementsaccompanyingmedrxivpreprintsandacomparisonwiththeirpublishedcounterparts |