ZICQ
中 Log in / Sign up
Newsroom LLMs & Foundation Models #ArXiv #Medical Imaging #Deep Learning #Ultrasound Imaging #Model Optimization

ArXiv Introduces AnatoProto Framework for Optimizing Standard Plane Detection in Fetal Ultrasound Blind Sweeps

Avatar of Mr.Xu

By Mr.Xu

Published:

中文阅读 (Chinese) English Version

Summary:The ArXiv team introduces AnatoProto, a lightweight sequence-level framework for detecting fetal abdominal circumference standard planes in low-cost obstetric blind sweeps. The framework leverages anatomy-weighted spatial pooling, within-case prototype loss, a three-stage cascade refinement process, and a hybrid prediction head to significantly enhance detection performance. On the ACOUSLIC-AI benchmark, AnatoProto achieves a test F1 score of 67.72, outperforming the strongest baseline by 15.76


Background and Challenge

Detecting standard planes, such as the fetal abdominal circumference standard plane, in fetal ultrasound blind sweeps is a highly imbalanced frame classification task. Positive frames account for less than 3% of the sequence and appear in short contiguous segments, posing significant challenges for traditional ultrasound and vision foundation models.

AnatoProto Framework

The ArXiv team introduces AnatoProto, a lightweight sequence-level framework that optimizes a frozen BiomedCLIP encoder through the following four key components:

  1. Anatomy-Weighted Spatial Pooling: Utilizes nnU-Net abdominal region probabilities as a spatial prior to reweight BiomedCLIP patch tokens, aggregating frozen semantic features onto anatomically meaningful regions.
  2. Within-Case Prototype Loss: Pulls each frame embedding toward the mean of positive frames in the same sweep, leveraging case-level structural information to compensate for the lack of frame-level information.
  3. Three-Stage Cascade Refinement: Elevates the prediction unit from noisy frames to structurally constrained segments, including frame->segment->case-level rejecters.
  4. Hybrid Prediction Head: Jointly models per-frame stability and inter-frame boundary transitions to suppress boundary false positives.

Experimental Results

On the ACOUSLIC-AI benchmark, AnatoProto achieves a test F1 score of 67.72, significantly outperforming the strongest baseline model (FetalCLIP + PRS, F1 = 54.52) by 13.20 percentage points and the video temporal-action-detection baseline (TriDet + PRS, F1 = 52.0) by 15.76 percentage points.

Technical Highlights

  • Anatomy-Weighted Spatial Pooling: Guides feature aggregation with spatial priors to enhance anatomical feature effectiveness.
  • Within-Case Prototype Loss: Utilizes case-level structural information to compensate for the lack of frame-level information.
  • Three-Stage Cascade Refinement: Optimizes from frame to segment to case level, enhancing prediction robustness.
  • Hybrid Prediction Head: Considers both frame stability and boundary transitions to suppress false positives.

Industry Impact and Developer Recommendations

AnatoProto provides an efficient solution for low-resource ultrasound image analysis, particularly in resource-constrained medical environments. Developers can draw inspiration from its multi-component collaborative optimization approach and apply it to other imbalanced classification tasks. Additionally, the anatomy-weighted spatial pooling and within-case prototype loss mechanisms have broad applicability in medical image analysis, automated diagnosis systems, and more.


Source: ArXiv cs.CV (2026-08-27)

— END —

Tags: #ArXiv #Medical Imaging #Deep Learning #Ultrasound Imaging #Model Optimization

Community Comments

Loading live comments and annotations…