Accuracy of Contrast-enhanced US for Differentiating Benign from Malignant Solid Small Renal Masses
Bibliographic record
Abstract
PURPOSE: To test the hypothesis that qualitative and quantitative features of contrast material-enhanced ultrasonography (US) can be used to differentiate benign from malignant small renal masses. MATERIALS AND METHODS: This is an institutional review board approved, HIPAA-compliant prospective study with written informed consent. Patients with histologically characterized solid small renal masses, excluding lipid-rich angiomyolipomas, underwent qualitative contrast-enhanced US with a combination of three different US machines. A subgroup of patients underwent quantitative contrast-enhanced US. Patients received a bolus injection of 0.2 mL of contrast material for qualitative and quantitative evaluations and were followed for 3 minutes. Two radiologists independently reviewed videotaped qualitative contrast-enhanced US examinations and were blinded to the final diagnoses. Features that were evaluated included lesion vascularity relative to the adjacent cortex in the arterial phase, the presence of a capsule, homogeneity, the pattern of vascularity, and washout. One radiologist separately reviewed a subset of contrast-enhanced US examinations that were performed with all three machines. Parameters of a first-pass time intensity curve were calculated for quantitative analysis. The Mann-Whitney test was used for quantitative parameters, the χ(2) or Fisher exact test was used for qualitative parameters, and κ statistics and Fleiss methodology were used to determine interobserver and intermachine agreement. RESULTS: The study population consisted of 91 patients (35 women and 56 men) with 94 lesions. The mean age was 62 years ± 14 (range, 21-91). Three patients had two lesions each, which were evaluated at two different sessions. There were 26 benign small renal masses (including 18 oncocytomas, seven lipid-poor angiomyolipomas, and one hemangioblastoma) and 68 malignant masses (including 41 clear cell, 20 papillary, and seven chromophobe renal cell carcinomas [RCCs[) that were 1.1-4.0 cm in diameter (mean, 2.7 cm ± 0.9). All patients underwent contrast-enhanced US on the same one machine, and 68 patients were imaged on all three machines. Vascularity was present in all lesions (n = 94) at contrast-enhanced US. Lesion hypovascularity relative to the adjacent cortex in the arterial phase was seen in only malignant lesions by both reviewers; reviewer 1 saw hypovascularity in 24 of 94 lesions (P = .0001), and reviewer 2 saw hypovascularity in 21 of 94 lesions (P = .0006), for a specificity of 100% (95% confidence interval [CI]: 84, 100). This feature had κ values of 0.91 (95%CI: 0.82, 1.00) between the two reviewers and 0.85 (95% CI: 0.72, 0.99) between the three machines. Eighteen of 20 papillary RCCs were hypovascular. Quantitative parameters of area under the receiver operating characteristics curve, peak intensity, wash-in slope of 10%-90% and 5%-45%, and washout slope of 100%-10% and 50%-10% were significantly higher in malignant renal masses (P = .018, P = .002, P = .036, P = .016, P = .001, and P = .005, respectively) than in benign lesions. CONCLUSION: Excluding lipid-rich angiomyolipoma, hypovascularity-which has high interobserver and intermachine agreement-of solid small renal masses relative to the cortex in the arterial phase has 100% specificity (95% CI: 84, 100) for detecting malignancy, most often papillary RCC.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".