MétaCan
Menu
Back to cohort
Record W2418925811

Bayesian Inference Methods Applied to Cancer Research

2009· dissertation· en· W2418925811 on OpenAlexaboutno aff
Rudy Gunawan

Bibliographic record

VenueUWSpace (University of Waterloo) · 2009
Typedissertation
Languageen
FieldBiochemistry, Genetics and Molecular Biology
TopicMolecular Biology Techniques and Applications
Canadian institutionsnot available
Fundersnot available
KeywordsInferenceBayesian inferenceBayesian probabilityComputer scienceMachine learningArtificial intelligence
DOInot available

Abstract

fetched live from OpenAlex

The purpose of this Thesis is to present a Bayesian analysis of oncological data sets with particular focus on cervical carcinomas and ovarian cancers. \n \nBayesian methods of data analysis have a very long history, and have been used with great success in many disciplines, from Physics to Econometrics. Nonetheless, they remain very controversial among statisticians who belong to the orthodox - i.e, frequentist school, and are not well known by the medical community. To help in that direction, we reviewed in the introductory chapter the basic philosophical and practical differences between the two schools, and in the second chapter, we briefly reviewed the history of Bayesian methodology, from the early efforts of Thomas Bayes and of Pierre Simon de Laplace to the modern contributions of Harold Jeffreys, Richard Cox, and Edwin Jaynes. \n \nIn many aspects of medical research, we deal with experimental data from which a certain proposition or hypothesis is validated. Unlike in physics, where we have strong and solid foundations such as Newton's law of motion, Snell's optical laws, Kirchoff's laws, Einstein's relativity theory, and many more, we do not have such privileges in medical research. Hence, many hypotheses are constantly tested as new evidence becomes available. One of the actively-researched medical areas is cancer research about which our understanding is still in its infancy. Numerous experiments (both in vivo and in vitro) and clinical trials have been conducted to further our knowledge; thus, Bayesian methodology finds its place to aid us in obtaining scientific inferences about certain propositions or hypotheses from available data and resources. \n \nIn this work, we use data given to us by our medical collaborators at the Princess Margaret Hospital (PMH) in Toronto to carry out two main projects: Firstly, to make an inference about the oxygenation status (oxygen partial pressure, pO$_2$) within human cervical carcinomas and secondly, an inference about the effectiveness of various molecularly-targeted agents (MTAs) in phase II clinical trials of relapsed ovarian cancer patients. \n \nIn the first problem, we address the challenges of tumor hypoxia - a state of oxygen deprivation in tumors. Currently, there are two methods to obtain tumor oxygen status, namely the direct Eppendorf needle electrode and the indirect immunohistochemical assay of a protein marker, Carbonic Anhydrase IX (CAIX). In this project, we introduce Bayesian probability theory to obtain inferences about tumor oxygenation from each technique and the concordance between the two techniques. From this study, we conclude that under certain conditions, two biopsies are sufficient to infer the tumor oxygenation level based on the immunohistochemical assays of CAIX. Additionally, there is a fair concordance between the direct and the indirect measurements of tumor oxygenation. \n \nIn the latter problem, ovarian cancer is the topic of study. Ovarian cancer has the highest mortality rate among gynecological cancers and one that is known to relapse. CA-125 is still the most inexpensive biomarker for monitoring ovarian cancers. From the phase II clinical trial data, we demonstrate the survival advantage of CA-125 responsive group of patients by means of a non-parametric Kaplan-Meier statistic.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.022
metaresearch head score (Gemma)0.054
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Theoretical or conceptual · Consensus signal: Theoretical or conceptual
GenreCandidate signal: Methods · Consensus signal: Methods
Teacher disagreement score0.022
Threshold uncertainty score0.116

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0220.054
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0020.002
Bibliometrics0.0050.006
Science and technology studies0.0010.005
Scholarly communication0.0040.003
Open science0.0020.003
Research integrity0.0030.006
Insufficient payload (model declined to judge)0.0060.002

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.023
GPT teacher head0.357
Teacher spread0.334 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designTheoretical or conceptual
Domainnot available
GenreMethods

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2009
Admission routes1
Has abstractyes

Explore more

Same venueUWSpace (University of Waterloo)Same topicMolecular Biology Techniques and ApplicationsFrench-language works237,207