Interim Positron Emission Tomography (PET) in Diffuse Large B-Cell Lymphoma: Independent Expert Nuclear Medicine Evaluation of ECOG 3404
Notice bibliographique
Résumé
Abstract Background: Positive interim PET scans have been associated with inferior outcomes in DLBCL treated with chemotherapy, alone or with rituximab. In the ECOG 3404 study for bulky and advanced DLBCL, PET scans at baseline and after 3 R-CHOP are centrally reviewed by a single reader; those with positive scans cross-over to R-ICE after 4 R-CHOP, whereas those with negative scans continue on R-CHOP. The primary endpoint of E3404 is progression-free survival. To determine the reproducibility of interim PET scan interpretation, we convened an expert panel. Methods: Three external nuclear medicine physicians visually scored baseline and interim PET scans independently and blinded to other clinical information or outcome. ECOG study criteria were binary (0,1) based on residual disease in initially involved sites with uptake greater than the liver. London criteria were on a scale of 0–5, where 4–5 was positive, based on increased uptake relative to the liver. Overall scores and agreement among experts were evaluated for both criteria, with application of the kappa statistic to correct for chance. Results: Using the ECOG criteria, external reviewers were in complete agreement in 68% of 38 interim scans and completely agreed with the central review in the same 68% cases. Agreement among the experts was 71% employing the London criteria in these cases. The range of PET+ interim scans by reviewer was 16.8% to 34.2% (p=NS) by both ECOG and London criteria. The kappa statistic for overall pairwise correlation between readers was 0.445 (0.396–0.533) using ECOG and 0.502 (0.396–0.630) using London criteria – indicating moderate consistency. Areas of disagreement often, but not exclusively, related to bone disease, the shape and focality of residual uptake, splenic disease, and rare scans without CT-fusion. Conclusions: These data show that visual criteria, either ECOG or London in this series, for interim PET results are moderately reproducible among individual nuclear medicine experts. Our finding of variability among experts indicates the need for caution in interpreting interim PET results in studies and in practice. Review of all E3404 cases is planned after accrual is completed (projected 12/08). Other ongoing studies evaluating interim PET after 1–3 cycles of therapy, the application of quantitative criteria, and consensus panels may provide further valuable information.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,030 | 0,036 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,000 |
| Méta-épidémiologie (sens large) | 0,001 | 0,001 |
| Bibliométrie | 0,003 | 0,001 |
| Études des sciences et des technologies | 0,001 | 0,001 |
| Communication savante | 0,001 | 0,001 |
| Science ouverte | 0,001 | 0,002 |
| Intégrité de la recherche | 0,001 | 0,000 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,002 | 0,001 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».