Pith. sign in

REVIEW 2 cited by

Tuning-Free Amodal Segmentation via the Occlusion-Free Bias of Inpainting Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2503.18947 v1 pith:PDUMO5PE submitted 2025-03-24 cs.CV

classification cs.CV
keywords segmentationamodalapproachinpaintingmasksmodelsbiasdatasets
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Amodal segmentation aims to predict segmentation masks for both the visible and occluded regions of an object. Most existing works formulate this as a supervised learning problem, requiring manually annotated amodal masks or synthetic training data. Consequently, their performance depends on the quality of the datasets, which often lack diversity and scale. This work introduces a tuning-free approach that repurposes pretrained diffusion-based inpainting models for amodal segmentation. Our approach is motivated by the "occlusion-free bias" of inpainting models, i.e., the inpainted objects tend to be complete objects without occlusions. Specifically, we reconstruct the occluded regions of an object via inpainting and then apply segmentation, all without additional training or fine-tuning. Experiments on five datasets demonstrate the generalizability and robustness of our approach. On average, our approach achieves 5.3% more accurate masks over the state-of-the-art.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. DeOcc-1-to-3: 3D De-Occlusion from a Single Image via Self-Supervised Multi-View Diffusion

    cs.CV 2025-06 conditional novelty 6.0 of 10

    A self-supervised fine-tuned multi-view diffusion model produces six consistent de-occluded views from one occluded image, improving downstream 3D reconstruction over two-stage baselines.

  2. GENA3D: Generative Amodal 3D Modeling by Bridging 2D Priors and 3D Coherence

    cs.CV 2025-11 conditional novelty 5.0 of 10

    A generative model reconstructs complete, occlusion-free 3D objects from sparse unposed views by combining 2D amodal inpainting with stereo-point-cloud-conditioned cross-attention.

Pith tools