MétaCan
Menu
Back to cohort
Record W4404409586 · doi:10.31219/osf.io/dsxzn

Can I have your data? Recommendations and practical tips for sharing neuroimaging data upon a direct personal request

2024· preprint· en· W4404409586 on OpenAlexaboutno aff
Anita S. Jwa, Martin Noergaard, Russell A. Poldrack

Bibliographic record

Venuenot available
Typepreprint
Languageen
FieldEnvironmental Science
TopicHealth, Environment, Cognitive Aging
Canadian institutionsnot available
Fundersnot available
KeywordsData sharingNeuroimagingComputer scienceInternet privacyData sciencePsychologyNeuroscienceMedicine

Abstract

fetched live from OpenAlex

Sharing neuroimaging data upon a direct personal request can be challenging both for researchers who request the data and for those who agree to share their data. Unlike sharing through repositories under standardized protocols and data use/sharing agreements, each party often needs to negotiate the terms of sharing and use of data case by case. This negotiation unfolds against a complex backdrop of ethical and regulatory requirements along with technical hurdles related to data transfer and management. These challenges can significantly delay the data sharing process, and if not properly addressed, lead to potential tensions and disputes between sharing parties. This study aims to help researchers navigate these challenges by examining what to consider during the process of data sharing and by offering recommendations and practical tips. We first divided the process of sharing data upon a direct personal request into six stages: requesting data, reviewing the applicability of and requirements under relevant laws and regulations, negotiating terms for sharing and use of data, preparing and transferring data, managing and analyzing data, and sharing the outcome of secondary analysis of data. For each stage, we identified factors to consider through a review of ethical principles for human subject research; individual institutions' and funding agencies' policies; and applicable regulations in the U.S. and E.U. We then provide practical insights from a large-scale ongoing neuroimaging data sharing project led by one of the authors as a case study. In this case study, PET/MRI data from a total of 782 subjects were collected through direct personal requests across seven sites in the USA, Canada, the UK, Denmark, Germany, and Austria. The case study also revealed that researchers should typically expect to spend an average of 8 months on data sharing efforts, with the timeline extending up to 24 months in some cases due to additional data requests or necessary corrections. The current state of data sharing via direct requests is far from ideal and presents significant challenges, particularly for early career scientists, who often have a limited time frame – typically two to three years – to work on a project. The best practices and practical tips offered in this study will help researchers streamline the process of sharing neuroimaging data while minimizing friction and frustrations.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.002
metaresearch head score (Gemma)0.001
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesMeta-epidemiology (narrow), Open science, Insufficient payload (model declined to judge)
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Simulation or modeling · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: none
Teacher disagreement score0.786
Threshold uncertainty score1.000

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0020.001
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0000.000
Science and technology studies0.0000.000
Scholarly communication0.0000.001
Open science0.0010.022
Research integrity0.0000.001
Insufficient payload (model declined to judge)0.0010.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.218
GPT teacher head0.411
Teacher spread0.193 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one teacher head, not a consensus.

Study designSimulation or modeling
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2024
Admission routes1
Has abstractyes

Explore more

Same topicHealth, Environment, Cognitive AgingFrench-language works237,207