MétaCan
Menu
Back to cohort
Record W3008436047 · doi:10.1371/journal.pone.0229182

The views, perspectives, and experiences of academic researchers with data sharing and reuse: A meta-synthesis

2020· article· en· W3008436047 on OpenAlexaff
Laure Perrier, Erik Blondal, H. Robson MacDonald

Bibliographic record

VenuePLoS ONE · 2020
Typearticle
Languageen
FieldComputer Science
TopicResearch Data Management Practices
Canadian institutionsCarleton UniversityOntario Council of University LibrariesUniversity of Toronto
Fundersnot available
KeywordsData sharingIncentiveData curationComputer scienceReuseData qualityData managementData scienceOpen dataThematic analysisKnowledge managementWorld Wide WebQualitative researchMedicineBusinessDatabaseSociologyEngineering

Abstract

fetched live from OpenAlex

BACKGROUND: Funding agencies and research journals are increasingly demanding that researchers share their data in public repositories. Despite these requirements, researchers still withhold data, refuse to share, and deposit data that lacks annotation. We conducted a meta-synthesis to examine the views, perspectives, and experiences of academic researchers on data sharing and reuse of research data. METHODS: We searched the published and unpublished literature for studies on data sharing by researchers in academic institutions. Two independent reviewers screened citations and abstracts, then full-text articles. Data abstraction was performed independently by two investigators. The abstracted data was read and reread in order to generate codes. Key concepts were identified and thematic analysis was used for data synthesis. RESULTS: We reviewed 2005 records and included 45 studies along with 3 companion reports. The studies were published between 2003 and 2018 and most were conducted in North America (60%) or Europe (17%). The four major themes that emerged were data integrity, responsible conduct of research, feasibility of sharing data, and value of sharing data. Researchers lack time, resources, and skills to effectively share their data in public repositories. Data quality is affected by this, along with subjective decisions around what is considered to be worth sharing. Deficits in infrastructure also impede the availability of research data. Incentives for sharing data are lacking. CONCLUSION: Researchers lack skills to share data in a manner that is efficient and effective. Improved infrastructure support would allow them to make data available quickly and seamlessly. The lack of incentives for sharing research data with regards to academic appointment, promotion, recognition, and rewards need to be addressed.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.220
metaresearch head score (Gemma)0.385
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch, Open science
Consensus categoriesMetaresearch
DomainCandidate signal: Reproducibility · Consensus signal: none
Study designCandidate signal: Qualitative · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: none
Teacher disagreement score0.996
Threshold uncertainty score0.962

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.2200.385
Meta-epidemiology (narrow)0.0020.002
Meta-epidemiology (broad)0.0050.007
Bibliometrics0.0230.026
Science and technology studies0.0040.008
Scholarly communication0.0140.020
Open science0.0040.008
Research integrity0.0030.004
Insufficient payload (model declined to judge)0.0040.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.697
GPT teacher head0.420
Teacher spread0.277 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; the direct Gemma label and the distilled Codex classifier agree on what is shown here.

Study designQualitative
DomainReproducibility
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations116
Published2020
Admission routes1
Has abstractyes

Explore more

Same venuePLoS ONESame topicResearch Data Management PracticesFrench-language works237,207