Data Management in Health-Related Research Involving Indigenous Communities in the United States and Canada: A Scoping Review
Bibliographic record
Abstract
Background: Multiple factors, including experiences with unethical research practices, have made some Indigenous groups in the United States and Canada reticent to participate in potentially beneficial health-related research. Yet, Indigenous peoples have also expressed a willingness to participate in research when certain conditions related to the components of data management—including data collection, analysis, security and storage, sharing, dissemination, and withdrawal—are met. A scoping review was conducted to better understand the terms of data management employed in health-related research involving Indigenous communities in the United States and Canada. Methods: PubMed, Embase, PsychINFO, and Web of Science were searched using terms related to the populations and topics of interest. Results were screened and articles deemed eligible for inclusion were extracted for content on data management, community engagement, and community-level research governance. Results: The search strategy returned 753 articles. 31 total articles were extracted, of which nine contained in-depth information on data management and underwent detailed extraction. All nine articles reported the development and implementation of data management tools, including research ethics codes, data-sharing agreements, and biobank access policies. These articles reported that communities were involved in activities and decisions related to data collection (n=7), data analysis (n=5), data-sharing (n=9), dissemination (n=7), withdrawal (n=4), and development of data management tools (n=9). The articles also reported that communities had full or shared ownership of (n=5), control over (n=9), access to (n=1), and possession of data (n=5). All nine articles discussed the role of community engagement in research and community-level research governance as means for aligning the terms of data management with the values, needs, and interests of communities. Conclusions: There is need for more research and improved reporting on data management in health-related research involving Indigenous peoples in the United States and Canada. Findings from this review can provide guidance for the identification of data management terms and practices that may be acceptable to Indigenous communities considering participation in health-related research.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.036 | 0.003 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.002 | 0.000 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.002 | 0.002 |
| Research integrity | 0.000 | 0.008 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".