Automated Assessment of Cardiac Systolic Function From Coronary Angiograms With Video-Based Artificial Intelligence Algorithms
Bibliographic record
Abstract
Importance: Understanding left ventricular ejection fraction (LVEF) during coronary angiography can assist in disease management. Objective: To develop an automated approach to predict LVEF from left coronary angiograms. Design, Setting, and Participants: This was a cross-sectional study with external validation using patient data from December 12, 2012, to December 31, 2019, from the University of California, San Francisco (UCSF). Data were randomly split into training, development, and test data sets. External validation data were obtained from the University of Ottawa Heart Institute. Included in the analysis were all patients 18 years or older who received a coronary angiogram and transthoracic echocardiogram (TTE) within 3 months before or 1 month after the angiogram. Exposure: A video-based deep neural network (DNN) called CathEF was used to discriminate (binary) reduced LVEF (≤40%) and to predict (continuous) LVEF percentage from standard angiogram videos of the left coronary artery. Guided class-discriminative gradient class activation mapping (GradCAM) was applied to visualize pixels in angiograms that contributed most to DNN LVEF prediction. Results: A total of 4042 adult angiograms with corresponding TTE LVEF from 3679 UCSF patients were included in the analysis. Mean (SD) patient age was 64.3 (13.3) years, and 2212 patients were male (65%). In the UCSF test data set (n = 813), the video-based DNN discriminated (binary) reduced LVEF (≤40%) with an area under the receiver operating characteristic curve (AUROC) of 0.911 (95% CI, 0.887-0.934); diagnostic odds ratio for reduced LVEF was 22.7 (95% CI, 14.0-37.0). DNN-predicted continuous LVEF had a mean absolute error (MAE) of 8.5% (95% CI, 8.1%-9.0%) compared with TTE LVEF. Although DNN-predicted continuous LVEF differed 5% or less compared with TTE LVEF in 38.0% (309 of 813) of test data set studies, differences greater than 15% were observed in 15.2% (124 of 813). In external validation (n = 776), video-based DNN discriminated (binary) reduced LVEF (≤40%) with an AUROC of 0.906 (95% CI, 0.881-0.931), and DNN-predicted continuous LVEF had an MAE of 7.0% (95% CI, 6.6%-7.4%). Video-based DNN tended to overestimate low LVEFs and underestimate high LVEFs. Video-based DNN performance was consistent across sex, body mass index, low estimated glomerular filtration rate (≤45), presence of acute coronary syndromes, obstructive coronary artery disease, and left ventricular hypertrophy. Conclusion and relevance: This cross-sectional study represents an early demonstration of estimating LVEF from standard angiogram videos of the left coronary artery using video-based DNNs. Further research can improve accuracy and reduce the variability of DNNs to maximize their clinical utility.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".