MétaCan
Menu
← Back to cohort

Abstract LB034: Analysis of PD-L1 IHC tests using NIST SRM 1934-traceable reference materials: A new paradigm for development of predictive IHC biomarkers

2021· article· en· W3180418576 on OpenAlexaff
Emina Torlakovic, Seshi R. Sompuram, Kodela Vani, Steve Bogen

Bibliographic record

VenueCancer Research · 2021
Typearticle
Languageen
FieldMedicine
TopicCancer Immunotherapy and Biomarkers
Canadian institutionsUniversity of SaskatchewanSaskatchewan Health Authority
Fundersnot available
KeywordsImmunohistochemistryGold standard (test)Detection limitNISTChemistryMedicinePathologyInternal medicineComputer scienceChromatography

Abstract

fetched live from OpenAlex

Abstract BACKGROUND. The challenges in accurate patient stratification for immune checkpoint inhibitors have been compounded by the fact that the FDA-cleared PD-L1 IHC tests are analytic ‘black boxes'. Relatively basic analytic parameters such as lower limit of detection and analytic dynamic range are unknown to both developers of assays as well as end users. The recent development of standardized PD-L1 immunohistochemistry (IHC) reference materials enables quantitative test characterizations that were not previously possible. METHODS. We surveyed 41 PD-L1 testing laboratories in North America and Europe, quantitatively defining each of the PD-L1 tests' analytic performance in terms of lower limit of detection and dynamic range. All four commercial PD-L1 kits were assessed by multiple laboratories. A variety of laboratory-developed tests (LDTs) we also assessed. The reference materials incorporated defined concentrations of PD-L1 peptide (intracellular domain) or recombinant extracellular domain protein, traceable to NIST Standard Reference Material 1934. Each laboratory received a slide with 10 separate PD-L1 calibrator concentrations: 2,200 - 600,000 molecules of PD-L1 extracellular domain or 34,000 - 2,200,000 molecules of PD-L1 intracellular domain. The calibrator concentrations ranged from those that are below the lower limit of detection to others that yield maximal staining. RESULTS. The data obtained with the four PD-L1 kits (VENTANA PD-L1 (SP263) Assay, VENTANA PD-L1 (SP142) Assay, DAKO PD-L1 IHC 28-8 pharmDx and DAKO PD-L1 IHC 22C3 pharmDx assays) revealed that the lower limits of detection (PD-L1 molecules per cell equivalent) are approximately: 50,000 - 180,000 (SP263), 800,000 - 1,200,000 (SP142), 220,000 - 360,000 (28-8), and 200,000 - 400,000 (22C3). The dynamic ranges for all of these tests are generally narrow, spanning less than a log concentration of PD-L1. The SP263 and SP142 assays showed no overlap of their analytic response curves. This means that a maximal stain intensity with SP263 kit can be associated with zero staining with the SP142 kit. Consequently, it is not possible to compensate for the variability in analytic sensitivity between these two tests by adjusting the percent positive cell cutoff. The 28-8 was more sensitive than, but statistically indistinguishable from the 22C3 assay. Laboratory-developed tests (LDTs) using these and other primary antibodies have their own unique analytic performance characteristics. CONCLUSIONS. The PD-L1 reference materials enable precise definitions of analytic test performance and linking them with clinical management thresholds. Therefore, this tool finds its most important implementation at the stage of development of new IHC predictive biomarkers in clinical trials as well as at the stage of methodology transfer to clinical IHC laboratories. Furthermore, our results also help define more precisely the possibility for assay interchangeability and to what degree the assays may be harmonized. Citation Format: Emina E. Torlakovic, Seshi Sompuram, Kodela Vani, Steve Bogen. Analysis of PD-L1 IHC tests using NIST SRM 1934-traceable reference materials: A new paradigm for development of predictive IHC biomarkers [abstract]. In: Proceedings of the American Association for Cancer Research Annual Meeting 2021; 2021 Apr 10-15 and May 17-21. Philadelphia (PA): AACR; Cancer Res 2021;81(13_Suppl):Abstract nr LB034.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.016
metaresearch head score (Gemma)0.013
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Bench or experimental · Consensus signal: Bench or experimental
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.016
Threshold uncertainty score0.086

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0160.013
Meta-epidemiology (narrow)0.0010.000
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0030.002
Science and technology studies0.0000.001
Scholarly communication0.0020.001
Open science0.0020.001
Research integrity0.0010.001
Insufficient payload (model declined to judge)0.0020.001

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.219
GPT teacher head0.463
Teacher spread0.244 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designBench or experimental
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2021
Admission routes1
Has abstractyes

Explore more

Same venueCancer Research→Same topicCancer Immunotherapy and Biomarkers→French-language works237,207→