REVIEW 16 cited by
AMOS: A Large-Scale Abdominal Multi-Organ Benchmark for Versatile Medical Image Segmentation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Despite the considerable progress in automatic abdominal multi-organ segmentation from CT/MRI scans in recent years, a comprehensive evaluation of the models' capabilities is hampered by the lack of a large-scale benchmark from diverse clinical scenarios. Constraint by the high cost of collecting and labeling 3D medical data, most of the deep learning models to date are driven by datasets with a limited number of organs of interest or samples, which still limits the power of modern deep models and makes it difficult to provide a fully comprehensive and fair estimate of various methods. To mitigate the limitations, we present AMOS, a large-scale, diverse, clinical dataset for abdominal organ segmentation. AMOS provides 500 CT and 100 MRI scans collected from multi-center, multi-vendor, multi-modality, multi-phase, multi-disease patients, each with voxel-level annotations of 15 abdominal organs, providing challenging examples and test-bed for studying robust segmentation algorithms under diverse targets and scenarios. We further benchmark several state-of-the-art medical segmentation models to evaluate the status of the existing methods on this new challenging dataset. We have made our datasets, benchmark servers, and baselines publicly available, and hope to inspire future research. Information can be found at https://amos22.grand-challenge.org.
Forward citations
Cited by 16 Pith papers
-
Learning Segmentation from Radiology Reports
R-Super converts tumor count, size, and location information from radiology reports into voxel-wise losses that improve CT tumor segmentation beyond training with masks alone.
-
HyperSORT: Self-Organising Robust Training with hyper-networks
A hyper-network that predicts segmentation UNet weights from per-sample learned latent vectors yields a structured map of annotation styles and a way to flag erroneous labels.
-
DrVD-Bench: Do Vision-Language Models Reason Like Human Doctors in Medical Image Diagnosis?
A new five-level medical imaging benchmark, DrVD-Bench, shows that vision-language models lose accuracy sharply as reasoning complexity grows and often diagnose without grounding in lesion evidence.
-
MultiverSeg: Scalable Interactive Segmentation of Biomedical Imaging Datasets with In-Context Guidance
MultiverSeg combines interactive prompting with a growing set of previously segmented image pairs to reduce the number of user interactions needed to segment a new biomedical dataset.
-
Curia-MAE: Multi-Modal Multi-Anatomy MAE Pre-Training for 3D Medical Image Segmentation
Curia-MAE is a multi-modal, multi-anatomy masked autoencoder whose frozen encoder modestly improves 3D segmentation over its MAE baseline, with the largest gains on lesion tasks.
-
SegMoTE: Token-Level Mixture of Experts for Medical Image Segmentation
SegMoTE shows that adding token-level mixture-of-experts routing to a frozen SAM decoder can match or beat medical-segmentation models trained on far more data, using 0.15M curated masks and 17M trainable parameters.
-
Large-scale Multi-sequence Pretraining for Generalizable MRI Analysis in Versatile Clinical Applications
A four-objective self-supervised pretraining recipe on a 336k-volume multi-sequence MRI corpus yields first-rank transfer on 39 of 44 downstream MRI tasks.
-
Is Visual in-Context Learning for Compositional Medical Tasks within Reach?
Training on synthetic compositional task sequences with sequence-level masking lets a transformer-based in-context learner follow multi-step medical imaging instructions on held-out images, but well below codebook upp...
-
RadSAM: Segmenting 3D radiological images with a 2D promptable model
RadSAM segments 3D organs in CT from a single point or box prompt by iteratively propagating the predicted mask as a prompt to neighboring slices, outperforming MedSAM and matching or beating nnU-Net on AMOS.
-
Leveraging Textual Anatomical Knowledge for Class-Imbalanced Semi-Supervised Multi-Organ Segmentation
Injecting GPT-4o-generated textual anatomical priors as segmentation-head parameters improves class-imbalanced semi-supervised multi-organ segmentation.
-
VOILA: Complexity-Aware Universal Segmentation of CT images by Voxel Interacting with Language
VOILA performs universal CT segmentation by contrastively aligning voxels with text prompts and training on complexity-graded samples, achieving competitive Dice scores with far fewer trainable parameters.
-
GIRAFE: Glottal Imaging Dataset for Advanced Segmentation, Analysis, and Facilitative Playbacks Evaluation
GIRAFE releases 65 color high-speed laryngeal videos, 760 manual glottal gap masks, automatic segmentation baselines, and facilitative playbacks.
-
Good Enough? An Investigation on the Impact of Label Quality in Large-Scale Medical Datasets
Label quality matters little when pre-training medical segmentation models, but still matters for in-domain deployment; only large quality gaps affect transfer results.
-
The Large Cancer Assistant (LCA): A Model-Agnostic Orchestration Framework for Scalable Clinical Decision Support in Oncology
A formal orchestration framework for oncology AI pipelines is proposed, demonstrating that routing logic and output schema remain invariant under model substitution via a proof-of-concept with synthetic stubs.
-
Diffusion-empowered AutoPrompt MedSAM
A class-index-driven diffusion-style prompt encoder turns MedSAM into a fully automatic segmenter that outputs semantically labeled masks, with reported gains on CT, MRI, endoscopy, and X-ray benchmarks.
-
A Unified Framework for Foreground and Anonymization Area Segmentation in CT and MRI Data
A nnU-Net-based toolkit segments body foreground and anonymized regions in 3D CT/MRI with high Dice scores.
Discussion (0). Continue with ORCID to comment.