Building a library of acute traumatic spinal cord injury images across Canada: a retrospective cohort study protocol
Bibliographic record
Abstract
INTRODUCTION: MRI is increasingly recognised as a valuable tool for assessing prognosis and predicting outcomes following traumatic spinal cord injury (SCI). Several potential MRI biomarkers have been identified, but efforts are still needed to improve the accuracy and feasibility of these biomarkers in clinical practice. This study aims to build a national Canadian SCI imaging repository for storing and analysing imaging data for SCI, with the goal of improving SCI MRI biomarkers to predict outcomes and inform clinical management. METHOD AND ANALYSIS: As a substudy of the Rick Hansen SCI Registry (RHSCIR), this retrospective multisite study includes individuals who sustained a traumatic cervical SCI between 2015 and 2021, were previously enrolled in RHSCIR, and had MRI scans acquired within 72 hours of injury and before any surgical intervention. Individuals with a penetrating trauma and/or with any prior spine surgery are excluded. The study principal investigator and research associates, experienced with data curation and with the standardised format and specifications of the Brain Imaging Data Structure standard, guide the site's curator on the steps to perform image deidentification and curation to create standardised datasets across all sites. These datasets are transferred to a Digital Research Alliance of Canada ('the Alliance') server designated for this project and concatenated to form the national Canadian SCI imaging repository (Neurogitea). We are using a semiautomated processing pipeline to quantify lesion morphology, together with additional imaging measures that are manually extracted from the images (for instance, the relative maximal spinal cord compression and the maximum canal compromise). Through linkage to RHSCIR clinical and epidemiological data already available on eligible participants, regression analysis is planned to predict neurological outcomes at discharge, including the American Spinal Injury Association Impairment Scale grade, upper and lower extremity motor and sensory scores. ETHICS AND DISSEMINATION: This protocol has been submitted by the participating sites to obtain ethics and institutional approvals prior to the study initiation at each site. All 12 sites across Canada have now obtained ethics and institutional approvals. Study results will be disseminated at local, national and international conferences and by journal publications.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.014 | 0.027 |
| Meta-epidemiology (narrow) | 0.002 | 0.002 |
| Meta-epidemiology (broad) | 0.002 | 0.002 |
| Bibliometrics | 0.009 | 0.010 |
| Science and technology studies | 0.008 | 0.003 |
| Scholarly communication | 0.005 | 0.003 |
| Open science | 0.005 | 0.003 |
| Research integrity | 0.002 | 0.002 |
| Insufficient payload (model declined to judge) | 0.028 | 0.007 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".