Right ventricular function is an important prognostic marker in pulmonary hypertension and other cardiovascular diseases, yet most echocardiography-focused artificial intelligence work has concentrated on the left ventricle. The authors developed a single-view deep-learning model called PH-ECHO-AI to provide a comprehensive right-heart–focused interpretation from a single apical four-chamber (A4C) echocardiographic clip. The model's declared goals were to produce four-chamber segmentation, localise landmarks, estimate biventricular ejection fraction (EF), quantify deformation and annular excursion, and predict confirmed PH from geometry derived from the single A4C view.
PH-ECHO-AI was trained using 8,416 A4C clips assembled from four publicly accessible datasets: EchoNet-Dynamic, CAMUS, RVENet, and MIMIC-IV-ECHO. No proprietary institutional imaging was used for development. A supervised head trained on 3D-echocardiographic RVEF labels was included to enable direct regression of RVEF without assuming RV geometric models. The authors indicate that code and trained model weights will be made available to credentialed investigators for independent evaluation after publication.
Model evaluation used held-out, training-excluded data with expert-reviewed reference standards and a per-cohort audit to preserve patient-level separation between development and test sets. Specific evaluation cohorts comprised 1,416 clips for segmentation performance assessment; 600 clips for biventricular function and deformation analysis (350 referenced to 3D-echocardiographic RVEF, 250 referenced to EchoNet LVEF); and 1,076 MIMIC-IV patients for PH prediction. PH prediction used five-fold cross-validation entirely within the external MIMIC-IV cohort. The authors note that RVEF evaluation was clip-disjoint but same-source, and therefore cross-centre RVEF validation remains outstanding.
Reported performance measures included Dice coefficient for segmentation, correlation coefficients (r), mean absolute error (MAE), Bland–Altman agreement for continuous estimates, and area under the receiver operating characteristic curve (AUC) for classification tasks. Calibration for PH prediction was assessed using the Brier score. Confidence intervals were reported for key correlation estimates where available.
Four-chamber segmentation generalized across the included datasets. Pooled Dice scores were reported as follows: LV 0.925, RV 0.836, LA 0.910, and RA 0.904. These values indicate strong overlap between model-predicted and reference masks for left-sided chambers and somewhat lower, but acceptable, performance for the more geometrically complex right ventricle.
For the left ventricle, single-view LVEF estimation achieved correlation r = 0.845 (95% CI 0.786 to 0.886) with a mean absolute error of 4.67% compared with reference LVEF.
For the right ventricle, the model regressed RVEF directly from the clip using 3D-echocardiographic supervision; this yielded r = 0.754 (95% CI 0.690 to 0.806) and MAE 4.98%. The authors state this performance matches published ceilings for single-view RVEF estimation and exceeds the correlation observed for geometric RV fractional area change (RVFAC), which was reported as r = 0.278 in their comparisons. The RVEF evaluation was performed on clip-disjoint, same-source data, so external, cross-centre RVEF validation was not completed in this work.
The model produced deformation and excursion metrics derived from the single A4C view, including RV free-wall and LV A4C longitudinal strain and tricuspid and mitral annular plane systolic excursion (TAPSE and MAPSE). The authors report these metrics were physiologically coherent in their analyses, indicating internal consistency with expected mechanical behaviour, though detailed numeric comparisons to external reference standards are not provided beyond the deformation description in the source.
PH prediction was developed and evaluated entirely within the MIMIC-IV cohort (1,076 patients) using echocardiographic geometry alone. The model detected confirmed PH with an area under the receiver operating characteristic curve (AUC) of 0.697 and showed strong calibration with a Brier score of 0.061. The authors emphasise that PH development and validation were cohort-contained to the MIMIC-IV dataset.
Key limitations noted by the authors include the absence of cross-centre RVEF validation because RVEF evaluation used same-source data that were clip-disjoint from training data. PH prediction was trained and tested entirely within a single external cohort, which may limit generalisability until external prospective validation is performed. The study used de-identified, publicly released or credentialed-access datasets under their respective data-use agreements and operated under institutional research ethics approval (University Health Network / CAPCR #26-5021) with a waiver of individual informed consent for retrospective analysis. The authors declared no competing interests. Code and model weights will be made available upon publication to credentialed investigators for independent evaluation, per the source.
PH-ECHO-AI is a single, reproducible deep-learning model that provides comprehensive right-heart–focused interpretation from one apical four-chamber view. It generalised segmentation across multiple datasets, produced competitive single-view RVEF accuracy comparable with dedicated RV models while exceeding geometric RVFAC correlation, estimated LVEF with high correlation and modest MAE, delivered physiologically coherent deformation and annular excursion metrics (TAPSE and MAPSE), and detected confirmed pulmonary hypertension with modest discrimination and good calibration within the MIMIC-IV cohort. The authors have committed to releasing code and trained weights for independent validation, while acknowledging remaining needs for cross-centre RVEF validation and broader external prospective testing.