Determining Cell-Of-Origin Subtypes In Diffuse Large B-Cell Lymphoma Using Gene Expression Profiling On Formalin-Fixed Paraffin-Embedded Tissue – An L.L.M.P.P. Project
Notice bibliographique
Résumé
Abstract The diffuse large B-cell lymphoma (DLBCL) cell-of-origin (COO) distinction into germinal center B cell (GCB) and activated B cell (ABC) subtypes, as molecularly described by our group, has profound biological, prognostic, and potential therapeutic implications. New therapeutic agents with selective activity in ABC and GCB DLBCL are under development. An accurate diagnostic assay is urgently needed to qualify patients for clinical trials using targeted agents and as a predictive biomarker. Although the subtypes were originally defined using gene expression profiling on snap-frozen tissues (frozen-GEP), it has become common practice to use less precise but relatively inexpensive and broadly applicable immunohistochemical (IHC) methods in formalin-fixed paraffin-embedded tissue (FFPET). We sought to create a robust, highly accurate molecular assay for COO distinction using new GEP techniques applicable to FFPET. Studies were performed on centrally reviewed DLBCL FFPET biopsies using cases that had “gold standard” COO assigned by frozen-GEP using Affymetrix U133 plus 2.0 microarrays. The training cohort consisted of 51 cases comprising 20 GCB, 19 ABC and 12 Unclassifiable (U) cases. An independent validation cohort, consisting of 68 cases (28 GCB, 30 ABC, 10 U) drawn from the validation cohort of Lenz et al (NEJM 2008) had the typical proportions of COO subtypes seen in DLBCL populations. Nucleic acids were extracted from 10um FFPET scrolls. Digital gene expression was performed on 200ng of RNA using NanoString technology (Seattle, WA). All FFPET GEP studies were performed in parallel at two independent sites (BC Cancer Agency, Vancouver, and NCI, Frederick, MD) using different FFPET scrolls to determine inter-site concordance and assess the robustness and portability of the assay. To assign COO by IHC, tissue microarrays were made using 0.6mm duplicate cores from 60/68 validation cohort cases, and stained for CD10, BCL6, MUM1, FOXP1, GCET1 and LMO2. Two hematopathologists independently assessed the proportion of tumor cells stained, with consensus on discordant cases reached with a third hematopathologist. For the validation studies, those producing and analyzing the GEP and IHC data were blinded to the “gold standard” COO. All 119 FFPET biopsies yielded sufficient RNA. A pilot study using the training cohort identified 20 genes (15 genes of interest and 5 house keeping genes) whose expression, measured using NanoString, would allow accurate replication of the COO assignment model of Lenz et al (NEJM 2008). NanoString was then used to quantitate these 20 genes in the training cohort, allowing the COO model to be optimized. Despite the age of the FFPET blocks (6-32 years old), 95% (49/51) of the training samples gave gene expression data of sufficient quality. The model, including coefficients, thresholds and QC parameters was then “locked” and applied to the independent validation cohort. Ninety-nine percent (67/68) of the samples from the validation cohort (5-12 years old) provided gene expression of adequate quality. Three cases did not give interpretable IHC results. When considering the “gold standard” ABC and GCB cases, the COO assignments by the NanoString assay at the NCI site were 93% concordant, with 5% labeled U and 1 ABC misclassified as GCB (see table). This 2% rate of misclassification of ABC and GCB cases compares favorably with the 9%, 6% and 17% rates for the interpretable cases from the Hans, Tally and Choi algorithms, respectively. Furthermore, the 98% concordance of COO assignment (95% if “gold standard” U cases are also included) between the NCI and BC Cancer Agency sites indicates that, in contrast to the IHC algorithms, the assay is reproducible.TableNanoString GEP assay - NCIHans algorithmTally algorithmChoi algorithmGCBUABCGCBNon-GCBGCBABCGCBABCFrozen GEPGCB2800210183192U721552864ABC1325422026620 In summary, 119 well-characterized DLBCL cases from the LLMPP, previously subtyped by our published disease-defining algorithm using frozen-GEP, were used to develop a highly accurate and robust NanoString 20 gene assay, applicable to RNA from FFPET that is routinely obtained for diagnosis. This new assay shows excellent performance in archival FFPET, and the rapid turn-around time (<36 hours from FFPET block to result) will allow prospective implementation in future therapeutic trials and, ultimately, clinical practice. Disclosures: No relevant conflicts of interest to declare.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,001 | 0,001 |
| Méta-épidémiologie (sens strict) | 0,000 | 0,000 |
| Méta-épidémiologie (sens large) | 0,000 | 0,000 |
| Bibliométrie | 0,001 | 0,000 |
| Études des sciences et des technologies | 0,000 | 0,000 |
| Communication savante | 0,000 | 0,000 |
| Science ouverte | 0,000 | 0,000 |
| Intégrité de la recherche | 0,000 | 0,000 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,001 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».