Promising microRNAs in pre-diagnostic serum associated with lung cancer up to eight years before diagnosis: a HUNT study
Bibliographic record
Abstract
INTRODUCTION: Blood biomarkers for early detection of lung cancer (LC) are in demand. There are few studies of the full microRNome in serum of asymptomatic subjects that later develop LC. Here we searched for novel microRNA biomarkers in blood from non-cancer, ever-smokers populations up to eight years before diagnosis. METHODS: Serum samples from 98,737 subjects from two prospective population studies, HUNT2 and HUNT3, were considered initially. Inclusion criteria for cases were: ever-smokers; no known cancer at study entrance; 0-8 years from blood sampling to LC diagnosis. Each future LC case had one control matched to sex, age at study entrance, pack-years, smoking cessation time, and similar HUNT Lung Cancer Model risk score. A total of 240 and 72 serum samples were included in the discovery (HUNT2) and validation (HUNT3) datasets, respectively, and analysed by next-generation sequencing. The validated serum microRNAs were also tested in two pre-diagnostic plasma datasets from the prospective population studies NOWAC (n = 266) and NSHDS (n = 258). A new model adding clinical variables was also developed and validated. RESULTS: Fifteen unique microRNAs were discovered and validated in the pre-diagnostic serum datasets when all cases were contrasted against all controls, all with AUC > 0.60. In combination as a 15-microRNAs signature, the AUC reached 0.708 (discovery) and 0.703 (validation). A non-small cell lung cancer signature of six microRNAs showed AUC 0.777 (discovery) and 0.806 (validation). Combined with clinical variables of the HUNT Lung Cancer Model (age, gender, pack-years, daily cough parts of the year, hours of indoor smoke exposure, quit time in years, number of cigarettes daily, body mass index (BMI)) the AUC reached 0.790 (discovery) and 0.833 (validation). These results could not be validated in the plasma samples. CONCLUSION: There were a few significantly differential expressed microRNAs in serum up to eight years before diagnosis. These promising microRNAs alone, in concert, or combined with clinical variables have the potential to serve as early diagnostic LC biomarkers. Plasma is not suitable for this analysis. Further validation in larger prospective serum datasets is needed.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".