cd /news/artificial-intelligence/open-ultrasound-foundation-model-for… · home topics artificial-intelligence article
[ARTICLE · art-133345] src=arxiv.org ↗ pub= topic=artificial-intelligence verified=true sentiment=↑ positive

Open ultrasound foundation model for robust segmentation and clinical measurement across heterogeneous settings

Researchers released SonoCorpus, an open ultrasound dataset of 456,963 images and 1,626,085 expert masks drawn from 53 public datasets across 24 clinical applications and 17 countries, alongside SonoBase, an interactive segmentation foundation model pretrained on it. SonoBase outperformed SAM2, MedSAM2, and MedSAM3 on all fifteen evaluation datasets, matched per-dataset specialist models, and recovered usable segmentations in 81% of cases where a baseline failed outright, including on handheld probes used by minimally trained operators in Sierra Leone and Tanzania. The team released all checkpoints, optimizer states, data-split indices, deduplication hashes, and starter code to support reproducibility.

by read1 min views1 publishedSep 18, 2026

arXiv:2609.19230v1 Announce Type: new Abstract: Ultrasound is the most widely deployed imaging modality worldwide, yet clinical AI remains fragmented into narrow single-task models that fail when device, operator, or anatomy changes. Here we present SonoCorpus, an open resource unifying 456,963 images and 1,626,085 expert masks from 53 public datasets spanning 24 clinical applications and 17 countries, and SonoBase, an interactive segmentation foundation model pretrained on it. Across fifteen evaluation datasets introducing new organs, devices, operators, and geographies, SonoBase outperforms SAM2, MedSAM2, and the concept-promptable MedSAM3 on every dataset and matches per-dataset specialist models trained on the same data; on fully external data it exceeds the accuracy these baselines achieve on their own in-distribution benchmarks. Ejection fraction derived from its segmentations falls within inter-observer variability (6.63% error), with fewer misclassifications at the defibrillator-candidacy threshold than either promptable baseline (13% versus 18--42%); fetal head-circumference (1.81~mm) and gestational-age (1.2 days) errors fall below inter-observer variability. Where a baseline fails outright, one in four test cases, SonoBase recovers a usable segmentation in 81% of them, including on handheld probes operated by minimally trained users in two low- and middle-income countries (Sierra Leone and Tanzania). Five labeled examples can help the model adapt to a new setting, and the identical training protocol transfers well to newer models such as SAM3, locating the advantage in ultrasound-specific pretraining rather than any single architecture. To ensure reproducibility and enable the community to build on SonoBase as a platform, we release all checkpoints, optimizer states, data-split indices, deduplication hashes, and starter code.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @sonocorpus 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/open-ultrasound-foun…] indexed:0 read:1min 2026-09-18 ·