Pith. sign in

REVIEW 5 cited by

$F$, $B$, Alpha Matting

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2003.07711 v1 pith:ETKRBZJM submitted 2020-03-17 cs.CV

classification cs.CV
keywords alphacoloursforegroundmattingnetworksbackgroundestimatingexisting
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Cutting out an object and estimating its opacity mask, known as image matting, is a key task in many image editing applications. Deep learning approaches have made significant progress by adapting the encoder-decoder architecture of segmentation networks. However, most of the existing networks only predict the alpha matte and post-processing methods must then be used to recover the original foreground and background colours in the transparent regions. Recently, two methods have shown improved results by also estimating the foreground colours, but at a significant computational and memory cost. In this paper, we propose a low-cost modification to alpha matting networks to also predict the foreground and background colours. We study variations of the training regime and explore a wide range of existing and novel loss functions for the joint prediction. Our method achieves the state of the art performance on the Adobe Composition-1k dataset for alpha matte and composite colour quality. It is also the current best performing method on the alphamatting.com online evaluation.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. LayeringDiff: Layered Image Synthesis via Generation, then Disassembly with Generative Knowledge

    cs.CV 2025-01 conditional novelty 6.5 of 10

    LayeringDiff synthesizes layered images by generating a composite with a pretrained diffusion model and then decomposing it into foreground and background layers using small fine-tuned networks.

  2. SDMatte: Grafting Diffusion Models for Interactive Matting

    cs.CV 2025-08 conditional novelty 6.0 of 10

    SDMatte adapts Stable Diffusion to interactive matting via visual-prompt cross-attention, opacity/coordinate embeddings, and masked self-attention, reporting SOTA results on multiple benchmarks.

  3. Memory Efficient Matting with Adaptive Token Routing

    cs.CV 2024-12 conditional novelty 6.0 of 10

    Adaptive token routing with a lightweight refinement branch lets a ViT matting model run on full-resolution high-res images at about 12% of the memory of the ViTMatte baseline with only a small accuracy drop on Compos...

  4. Morpho-Aware Global Attention for Image Matting

    cs.CV 2024-11 conditional novelty 6.0 of 10

    A vision transformer matting model with Tetris-like convolutional query embeddings achieves top results on Composition-1k and Distinctions-646.

  5. BiVM: Accurate Binarized Neural Network for Efficient Video Matting

    cs.CV 2025-07 conditional novelty 5.0 of 10

    BiVM is a 1-bit binarized video matting network that beats prior binarized methods on accuracy and efficiency, with 11.82 MAD on VideoMatte240K versus 28.49 for ReActNet-binarized RVM.

Pith tools