Skip to content
Open access

Multi-Task Deep Learning Framework for Real-Time Quality Assessment and Probe Guidance in Echocardiography.

Jul 2026 · IEEE journal of biomedical and health informatics · Vol PP, pp. 1-15 · 0 citations
Medicine

TL;DR

A multi-task guidance framework that jointly performs supervised view classification and cardiac structure segmentation and reuses their outputs to enable label-efficient, structure-specific image quality scoring through entropy-based metrics without additional quality annotations can assist novice or trainee users in consistently acquiring acceptable echocardiographic views with minimal additional annotation.

Abstract

Transthoracic echocardiography is a widely used, noninvasive tool for cardiac imaging, but the quality of image acquisition remains highly operator-dependent. Existing artificial intelligence-based systems often require large amounts of labeled data and provide only global quality scores, limiting their utility in real-time clinical applications. We propose a multi-task guidance framework that jointly performs supervised view classification and cardiac structure segmentation, and reuses their outputs to enable label-efficient, structure-specific image quality scoring through entropy-based metrics without additional quality annotations. A lightweight maneuver predictor then uses these features to suggest one of seven corrective probe maneuvers in real-time. To train and validate the system, we constructed a 43-case maneuver-tagged dataset capturing intentional transitions from standard to nonstandard views. The proposed quality metric successfully distinguished standard from nonstandard views across multiple cardiac structures (AUC: 0.901-0.987). The maneuver predictor achieved a top-1 accuracy of 85.1% (mAP: 0.904) and inference time of 17 ms per frame, supporting its feasibility for real-time use. This system can assist novice or trainee users in consistently acquiring acceptable echocardiographic views with minimal additional annotation, which can improve clinical efficiency and reliability.

Read PDF

Similar papers

Aug 2026

Structure-Aware Deep Learning for Pediatric Echocardiographic Standard-View Classification in Ultrasonic Imaging.

Accurate standard-view classification is essential for pediatric echocardiographic image analysis and downstream automated interpretation. This task remains challenging because discriminative view information is often encoded in subtle chamber configurations, outflow-tract morphology, and weak anatomical boundaries, whereas conventional classifiers may underuse shallow and intermediate representations that preserve spatial structure. We propose PVTv2-ASEF, a structure-aware framework for 4-class pediatric echocardiographic standard-view classification. The framework introduces an Adaptive Structural Enhancement Module that performs residual input-side conditioning through channel recalibration, local convolutional mixing, multi-scale structural modeling, and input-dependent branch weighting. It further employs a Dual Auxiliary Fusion Head to transform Stage 2 and Stage 3 representations into class-level evidence and fuse them with the final-stage logits during inference. PVTv2-ASEF was evaluated on a private pediatric ventricular septal defect echocardiography dataset comprising 4 standard views under repeated patient-disjoint evaluation, with macro-F1 used as the primary class-balanced metric. Compared with PVTv2-B2, the proposed framework improved macro-F1 by 0.093 on the private dataset and by 0.023 in cross-task evaluation on FETAL_PLANES_DB. These results support the utility of coordinated input-side structural enhancement and intermediate logit fusion for ultrasound view and plane classification.

Yanfeng Liu, Haibin Sun, Hai-Song Huang et al. · 0 citations
Open access Jul 2026

BackMix-Enhanced Semi-Supervised Learning for Automated Detection of Aortic Stenosis from Transthoracic Echocardiographic Images

The proposed anatomically guided BackMix augmentation combined with semi-supervised ensemble learning can improve classification accuracy, robustness, and interpretability in echocardiographic analysis under limited annotation conditions, offering a promising approach for automated AS assessment across independent clinical datasets.

Fatima Ezzahra Elkouahy, Badreddine Labakoum, H. Ouahid et al. · 0 citations
Jul 2026

CARDIAG: A Dense Segment Classification Benchmark of Deep Learning Architectures for Coronary Angiography

Accurate pixel-level classification of coronary angiograms is critical for cardiovascular disease assessment, yet the field lacks standardized evaluation protocols. In this work we demonstrate a new benchmark for the assessment of deep learning models which densely classify pixels of coronary angiograms to one of SYNTAX classes (or background). The evaluation covers 24 distinct architectures starting with classic convnets to recent state-space-based vision algorithms. We release CARDIAG - a multi-center, multi-label dataset which we carefully split to reliably compute metrics, accounting for diameter error, overlap, centerline quality and calibration. The data contains SYNTAX labels, binary, uncertainty and segmentation masks as well as intermediate frames together with the selected non-sensitive DICOM metadata. From the multitude of algorithms, we nominate ConvNeXt V2 encoder with DeepLab V3 Plus decoder as the best performing, achieving macro $F_1=0.456$, which we then ensemble with Mamba U-Net and Feature Pyramid Network, for an increased $F_1=0.479$. We demonstrate all the architectures to be well calibrated and determine the generalization of the top 5 methods, together with the data efficiency of these architectures. We highlight the importance of both high-resolution and low-resolution features in encoding. We also demonstrate the model correctness in the context of patient demographic, vessel sides and projection angle configurations. Overall the released benchmark allows for future studies to robustly and rigorously assess the proposals, not only for SYNTAX segmentation, but lesion detection and many more.

Dominik Bernard Lau, Hubert Malinowski, Jerzy Szyjut et al. · 0 citations
#artificial intelligence Preprint Aug 2026

MR-JEPA: A General Purpose Video Foundation Model for Cardiac MRI

Cardiac magnetic resonance imaging (CMR) produces rich sequential data such as temporal cine videos and spatial LGE/mapping stacks, yet most deep learning approaches process individual 2D slices, discarding this context. We present MR-JEPA, a self-supervised video foundation model for CMR that extends LeJEPA to 3D spatiotemporal inputs through tubelet tokenization, spatiotemporal masking augmentation, and initialization from a 2D CMR foundation model. Unlike prior CMR video models limited to cine data, MR-JEPA is pretrained on multi-sequence data (cine, LGE, mapping) from 10,505 patients across two centers without annotations. We evaluate the frozen encoder on six downstream tasks using a unified multi-view gated attention architecture: LV ejection fraction, RV ejection fraction, three myocardial strains (GLS, GCS, GRS), and four-class disease detection. MR-JEPA outperforms other compared methods on all five regression tasks, including both a domain-specific CMR model pretrained on more data with text supervision and a natural-video foundation model, achieving an LV EF MAE of 4.79% (r =0.764) and a GLS MAE of 1.87 (r=0.805), with 21-27% MAE reductions over baselines on strain tasks. For disease detection, MR-JEPA achieved a macro AUG of 0.868, remaining competitive with the domain-specific baseline despite using a fully self-supervised pretraining objective. These results demonstrate the potential of a unified video encoder for robust, multi-view utilization of diverse CMR sequences in clinical cardiac quantification and diagnosis.

Athira J. Jacob, Puneet Sharma, D. Comaniciu et al. · 0 citations
Open access Jul 2026

A deep learning framework for standardized interpretation of multiparameter cardiac ultrasound and disease classification

DeepCard is a multi-task deep learning system that produces standardized, reproducible interpretation of pre-measured echocardiographic parameters by jointly analyzing 39 quantitative measurements across 17 diagnostic tasks spanning valvular disease, ventricular dysfunction, and structural abnormalities.

Zhihong Wen, Xiang-Peng Liu, Yi Liu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.