Outcomes of Measurable Residual Disease in Pediatric Acute Myeloid Leukemia before and after Hematopoietic Stem Cell Transplant: Validation of Difference from Normal Flow Cytometry with Chimerism Studies and Wilms Tumor 1 Gene Expression
Bibliographic record
Abstract
We enrolled 150 patients in a prospective multicenter study of children with acute myeloid leukemia undergoing hematopoietic stem cell transplantation (HSCT) to compare the detection of measurable residual disease (MRD) by a "difference from normal" flow cytometry (ΔN) approach with assessment of Wilms tumor 1 (WT1) gene expression without access to the diagnostic specimen. Prospective analysis of the specimens using this approach showed that 23% of patients screened for HSCT had detectable residual disease by ΔN (.04% to 53%). Of those patients who proceeded to transplant as being in morphologic remission, 10 had detectable disease (.04% to 14%) by ΔN. The disease-free survival of this group was 10% (0 to 35%) compared with 55% (46% to 64%, P < .001) for those without disease. The ΔN assay was validated using the post-HSCT specimen by sorting abnormal or suspicious cells to confirm recipient or donor origin by chimerism studies. All 15 patients who had confirmation of tumor detection relapsed, whereas the 2 patients with suspicious phenotype cells lacking this confirmation did not. The phenotype of the relapse specimen was then used retrospectively to assess the pre-HSCT specimen, allowing identification of additional samples with low levels of MRD involvement that were previously undetected. Quantitative assessment of WT1 gene expression was not predictive of relapse or other outcomes in either pre- or post-transplant specimens. MRD detected by ΔN was highly specific, but did not identify most relapsing patients. The application of the assay was limited by poor quality among one-third of the specimens and lack of a diagnostic phenotype for comparison.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".