Searcharxiv⌕ Search

arXiv · 2609.31376

Towards Whole-Study Screening for Congenital Heart Disease in Fetal Ultrasound Using Multiple Instance Learning

Abstract

Congenital heart disease (CHD) is the most common birth defect, yet a large fraction of cases remain undetected on prenatal ultrasound, in part because current artificial-intelligence methods assume that the key diagnostic frames have already been isolated from a study, by a clinician or by a view classifier. We remove that assumption and address CHD screening directly at the level of the whole ultrasound study. We propose a two-stage framework that first learns transferable frame representations by self-supervised masked-autoencoder pre-training on unlabeled fetal ultrasound, then identifies cardiac frames with a disease-robust module and aggregates them with a transformer-based multiple instance learning (MIL) model that produces a case-level diagnosis from study-level labels alone. The model further returns its highest-scoring frames for clinician review, and a hierarchical head separates critical from non-critical CHD. On the internal test set of our multi-source development cohort (FUSE), the proposed cardiac-gated MIL model reaches an area under the curve (AUC) of 0.985 with a specificity of 0.990, outperforming the reproduced NATMED ensemble (AUC 0.861, specificity 0.600) and the FetalCLIP foundation model (AUC 0.867, specificity 0.710). On an independent external cohort, all models initially perform near chance, but label-free CORAL adaptation raises the proposed model from an AUC of 0.513 to 0.944, whereas whole-study and view-dependent baselines do not recover. These results indicate that whole-study MIL with disease-robust cardiac-frame identification is an accurate and deployable route to prenatal CHD screening.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Mohamed Azzam, Ruobing Liu, Esther C. Ugwueke, Ziyang Xu, Shibiao Wan, Alex Foy, Abraham Zabih, Jason Christensen, Neil Hamill, Ling Li, Jieqiong Wang. 2026-09-25. Towards Whole-Study Screening for Congenital Heart Disease in Fetal Ultrasound Using Multiple Instance Learning. https://arxiv.org/abs/2609.31376

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Distribution-Aligned Representation Adaptation for DJSCC over Hybrid Wireless-Wired Networks

Deep joint source-channel coding (DJSCC) has emerged as a robust alternative to traditional separate coding for communications through wireless channels. Existing DJSCC approaches focus primarily on point-to-point wireless communication scenarios, while neglecting end-to-end communication efficiency in hybrid wireless-wired networks such as 5G and 6G communication systems. Considerable redundancy in DJSCC symbols against wireless channels becomes inefficient for long-distance wired transmission. Furthermore, DJSCC symbols must adapt to the varying transmission rate of the wired network to avoid congestion. In this paper, we propose a novel framework designed for efficient wired transmission of DJSCC symbols within hybrid wireless-wired networks, namely Rate-Controllable Wired Adaptor (RCWA). RCWA achieves redundancy-aware coding for DJSCC symbols to improve transmission efficiency, which removes considerable redundancy present in DJSCC symbols for wireless channels and encodes only source-relevant information into bits. Moreover, we leverage the Lagrangian multiplier method to achieve controllable and continuous variable-rate coding, which can encode given features into expected rates, thereby minimizing end-to-end distortion while satisfying given constraints. Extensive experiments on diverse datasets demonstrate the superior RD performance and robustness of RCWA compared to existing baselines, validating its potential for wired resource utilization in hybrid transmission scenarios. Specifically, our method can obtain peak signal-to-noise ratio gain of up to 0.7dB and 4dB compared to neural network-based methods and digital baselines on CIFAR-10 dataset, respectively.

eess.IV↗

SCI-Mamba: Unsupervised Learning based Low-Light Image Enhancement for Non-Cooperative Spacecraft

Low-light visual perception acts as the core visual foundation for on-orbit servicing missions targeting non-cooperative spacecraft, supporting autonomous rendezvous, pose estimation, component detection and robotic capture operations. Spaceborne imagery suffers from severe low-light degradation, while the extreme scarcity of paired normal/low-light space samples severely limits the generalization capacity of supervised enhancement algorithms. To address this practical bottleneck, this paper proposes SCI-Mamba, an unsupervised enhancement network for low-light orbital spacecraft observations. The proposed framework unites self-calibrated unsupervised learning, linear-complexity VMamba architecture and Retinex physical priors, delivering a lightweight enhancement pipeline adaptable to resource-limited spaceborne hardware. We construct Space Dark-1.0, a dedicated low-light spacecraft dataset integrating real orbital footage, darkroom hardware-in-the-loop measurements and physically constrained synthetic data covering diverse illumination, motion and attitude conditions. Comprehensive comparisons with CNN-, Transformer- and prevailing Mamba-based approaches verify the advantages of SCI-Mamba in visual authenticity, color fidelity and inference speed. The proposed framework provides a practical low-light enhancement solution for close-proximity non-cooperative space operations. The code is available at https://github.com/bitswh/SCI-Mamba

eess.IV↗

Reliability assessment and multicenter clinical application of magnetic resonance methods for knee cartilage quantification

Background: This study evaluated interreader agreement and longitudinal performance of MRI methods for knee cartilage volume, thickness, and defect-area quantification. Methods: AI-presegmented masks from 1,189 phase III examinations underwent independent correction by two readers and adjudication. Cartilage volume, three-dimensional ray-tracing thickness (3D-RT), and ray-based defect area (3D-RBA), defined by a 1.5-mm thickness threshold, were calculated. Agreement was assessed using segmentation metrics, intraclass correlation coefficients (ICCs), repeated-measures Bland-Altman analysis, and minimal detectable change at 95% confidence (MDC95). The 3D-RBA framework was evaluated in 120 digital-phantom experiments from 40 participants. Longitudinal analyses included 374 participants, alternative-method comparisons included 65, and retrospective phase II analysis included 24 participants with four visits. Results: Overall AI-to-adjudicated-mask Dice was 0.964 +/- 0.029. Interreader ICCs for volume, thickness, and defect area were 0.956, 0.904, and 0.932; corresponding MDC95 values were 1,596.9 mm^3, 0.227 mm, and 147.4 mm^2. Geometric mean absolute percentage error for defect area was 5.62%, with spatial Dice of 0.961. In 374 participants, volume changes correlated positively with thickness changes (rho=0.431) and negatively with defect-area changes (rho=-0.221). Within-participant phase II correlations followed the same directions in both groups. Conclusions: The workflow demonstrated good interreader agreement. Controlled geometric results and longitudinal associations supported the feasibility of threshold-based defect-area estimation. Volume, thickness, and defect area provide complementary measures of cartilage structure.

eess.IV↗