MétaCan
Menu
Back to cohort

Abstract A023: Determining genetic interaction from double knockout CRISPR screening

2024· article· en· W4399505367 on OpenAlexaboutno aff
John Paul Shen, Yue Gu, Saikat Chowdhury

Bibliographic record

VenueMolecular Cancer Therapeutics · 2024
Typearticle
Languageen
FieldBiochemistry, Genetics and Molecular Biology
TopicBioinformatics and Genomic Networks
Canadian institutionsnot available
Fundersnot available
KeywordsCRISPRComputational biologyGeneticsHeLaGeneBiologyGenetic analysisGene knockoutCell

Abstract

fetched live from OpenAlex

Abstract Background For decades double knockout (KO) perturbation screens were limited to model organisms such as S. pombe and S. cerevisiae. CRISPR technology has revolutionized genetic interaction discovery by allowing large scale screening in human cell lines, organoids, and mouse models. However, there remains much uncertainty regarding the optimal way to determine the presence of genetic interaction from the raw data generated from these large scale double perturbation experiments. Here we compare two different analysis methods run on the same normalized dataset to determine to what degree does the analysis method influence the determination of genetic interaction. Methods A publicly available genetic interaction dataset containing 24,908 double knock-out constructs across three cell lines (Hela, A549, 293T) in four time points (day 3, 14, 21, 28) and two replicates generated from a pair-wise CRISPR-Cas9 KO screen was used for analysis (Shen et al, Nature Methods, 2017). These data were used to measure single gene fitness scores for 73 known cancer driver genes and all 2628 pair-wise interactions using (1) the numerical Bayesian method from Shen et al, called CTG (Compositional and Time-course-aware Genetic analysis), and (2) the variational Bayesian method GEMINI (Zamanighomi et al, Genome Biology, 2019). Results Single gene KO fitness measurements from CTG and GEMINI were highly correlated for all three cell lines (pearson r 0.678, 0.604, 0.784 for HeLa, A549, and 293T, respectively; p< 0.1 x10-8 for each). In contrast, correlation of genetic interaction scores between the two methods was essentially random: HeLa r= -0.0143, p= 0.47, A549; A549 r= -0.0476, p=0.015; 293T r= -0.0135, p= 0.49. Of 52 synthetic lethal interactions identified by CTG in HeLa at z-score cut off -3, none were identified by GEMINI at same Z cutoff. Conversely of 4 interactions identified by GEMINI, none were identified by CTG. Similarly in A549, of 57 interactions identified by CTG none were identified by GEMINI, of 3 interactions identified by GEMINI none were identified by CTG. Restricting to genetic interactions that were validated in low-throughput drug-drug assays, of 5 synthetic lethal interactions found in HeLa by CTG (CHEK1-MAP2K1, CHEK1-TYMS, ADA-CHEK1, ATM-CHEK1, CDK9-CHEK1) all but CHEK1-TYMS were validated in low-throughput assays. However none of the 5 were scored as hits by GEMINI. Of 3 interactions scored as synthetic lethal in A549 (PRKDC-RRM2, CDK9-PRKDC, CDK4-PRKDC) all but PRKDC-RRM2 were validated, none of the 3 were scored as hits by GEMINI. Conclusions This study highlights dramatic differences in calculated genetic interaction scores from two different computational algorithms applied to the same experimental data. With only 8 of 2628 (0.3%) interactions tested in validation experiments it is not currently possible to know the ground truth in order to assess which method is most accurate. The generation of synthetic genetic interaction data will be an important step for further optimization of algorithms to detect genetic interaction. Citation Format: John Paul Shen, Yue Gu, Saikat Chowdhury. Determining genetic interaction from double knockout CRISPR screening [abstract]. In: Proceedings of the AACR Special Conference in Cancer Research: Expanding and Translating Cancer Synthetic Vulnerabilities; 2024 Jun 10-13; Montreal, Quebec, Canada. Philadelphia (PA): AACR; Mol Cancer Ther 2024;23(6 Suppl):Abstract nr A023.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.003
metaresearch head score (Gemma)0.006
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Bench or experimental · Consensus signal: Bench or experimental
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.003
Threshold uncertainty score0.017

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0030.006
Meta-epidemiology (narrow)0.0010.000
Meta-epidemiology (broad)0.0010.001
Bibliometrics0.0020.001
Science and technology studies0.0010.000
Scholarly communication0.0010.001
Open science0.0010.001
Research integrity0.0010.001
Insufficient payload (model declined to judge)0.0030.002

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.028
GPT teacher head0.305
Teacher spread0.277 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designBench or experimental
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2024
Admission routes1
Has abstractyes

Explore more

Same venueMolecular Cancer TherapeuticsSame topicBioinformatics and Genomic NetworksFrench-language works237,207