Development of a Web-Based System for Exploring Cancer Risk With Long-term Use of Drugs: Logistic Regression Approach
Bibliographic record
Abstract
BACKGROUND: Existing epidemiological evidence regarding the association between the long-term use of drugs and cancer risk remains controversial. OBJECTIVE: We aimed to have a comprehensive view of the cancer risk of the long-term use of drugs. METHODS: A nationwide population-based, nested, case-control study was conducted within the National Health Insurance Research Database sample cohort of 1999 to 2013 in Taiwan. We identified cases in adults aged 20 years and older who were receiving treatment for at least two months before the index date. We randomly selected control patients from the patients without a cancer diagnosis during the 15 years (1999-2013) of the study period. Case and control patients were matched 1:4 based on age, sex, and visit date. Conditional logistic regression was used to estimate the association between drug exposure and cancer risk by adjusting potential confounders such as drugs and comorbidities. RESULTS: There were 79,245 cancer cases and 316,980 matched controls included in this study. Of the 45,368 associations, there were 2419, 1302, 662, and 366 associations found statistically significant at a level of P<.05, P<.01, P<.001, and P<.0001, respectively. Benzodiazepine derivatives were associated with an increased risk of brain cancer (adjusted odds ratio [AOR] 1.379, 95% CI 1.138-1.670; P=.001). Statins were associated with a reduced risk of liver cancer (AOR 0.470, 95% CI 0.426-0.517; P<.0001) and gastric cancer (AOR 0.781, 95% CI 0.678-0.900; P<.001). Our web-based system, which collected comprehensive data of associations, contained 2 domains: (1) the drug and cancer association page and (2) the overview page. CONCLUSIONS: Our web-based system provides an overview of comprehensive quantified data of drug-cancer associations. With all the quantified data visualized, the system is expected to facilitate further research on cancer risk and prevention, potentially serving as a stepping-stone to consulting and exploring associations between the long-term use of drugs and cancer risk.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".