REVIEW 18 cited by
Enhanced-alignment Measure for Binary Foreground Map Evaluation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The existing binary foreground map (FM) measures to address various types of errors in either pixel-wise or structural ways. These measures consider pixel-level match or image-level information independently, while cognitive vision studies have shown that human vision is highly sensitive to both global information and local details in scenes. In this paper, we take a detailed look at current binary FM evaluation measures and propose a novel and effective E-measure (Enhanced-alignment measure). Our measure combines local pixel values with the image-level mean value in one term, jointly capturing image-level statistics and local pixel matching information. We demonstrate the superiority of our measure over the available measures on 4 popular datasets via 5 meta-measures, including ranking models for applications, demoting generic, random Gaussian noise maps, ground-truth switch, as well as human judgments. We find large improvements in almost all the meta-measures. For instance, in terms of application ranking, we observe improvementrangingfrom9.08% to 19.65% compared with other popular measures.
Forward citations
Cited by 18 Pith papers
-
S3OD: Towards Generalizable Salient Object Detection with Synthetic Data
A 139k-image synthetic dataset with diffusion- and DINO-derived masks, trained with a multi-mask decoder, improves cross-dataset salient-object detection and reaches state-of-the-art after fine-tuning.
-
Is There Really a Camouflaged Object? Towards Realistic Camouflaged Object Detection
The authors propose a 16,245-image benchmark that includes negative samples for camouflaged object detection and a network that jointly predicts object presence, camouflage presence, and segmentation masks.
-
To Blend In, First Decouple: Rethinking Camouflage Image Generation via Context-Decoupled Representations
CamoDreamer generates camouflage images by decoupling foreground and background control in a diffusion model, reporting a 15.5-point FID gain over prior state of the art on LAKE-RED.
-
Mamba Guided Boundary Prior Matters: A New Perspective for Generalized Polyp Segmentation
SAM-MaGuP, a SAM-based polyp segmentation model with a 1D-2D Mamba adapter and boundary distillation, reports state-of-the-art mDice/mIoU on five public colonoscopy datasets.
-
TransGUNet: Transformer Meets Graph-based Skip Connection for Medical Image Segmentation
A graph-based skip connection with entropy-filtered spatial attention improves average medical-image segmentation accuracy and efficiency across 14 datasets.
-
Seamless Detection: Unifying Salient Object Detection and Camouflaged Object Detection
A task-agnostic network unifies salient and camouflaged object detection via contrastive foreground-background distillation, reaching competitive supervised and SOTA unsupervised results.
-
Concept Guided Co-salient Object Detection
ConceptCoSOD extracts a shared text embedding from an image group and guides co-salient object segmentation with it, outperforming five baselines on three clean and five corrupted datasets.
-
MSRNet: A Multi-Scale Recursive Network for Camouflaged Object Detection
MSRNet, a multi-scale recursive network with attention-based scale integration and recursive-feedback decoding, reports state-of-the-art or runner-up camouflaged object detection on four standard benchmarks.
-
Seg-R1: Segmentation Can Be Surprisingly Simple with Reinforcement Learning
Reinforcement learning can teach an LMM to prompt SAM2 for segmentation, achieving competitive camouflaged and salient object detection and zero-shot referring segmentation.
-
Learning Motion and Temporal Cues for Unsupervised Video Object Segmentation
MTNet fuses appearance and motion features with a mixed local-global temporal transformer to achieve state-of-the-art unsupervised video object segmentation results on DAVIS-16, FBMS, YouTube-Objects, and Long-Videos.
-
Boosting Salient Object Detection with Knowledge Distillated from Large Foundation Models
A text-driven BLIP-GroundingDINO-SAM pipeline produces pseudo-labels and a new 260k-image dataset for SOD, with claimed SOTA results that are weakened by a likely PASCAL-S overlap and missing artifacts.
-
DefFiller: Mask-Conditioned Diffusion for Salient Steel Surface Defect Generation
DefFiller fine-tunes GLIGEN with a mask encoder to synthesize steel defects that match given masks, and shows small but consistent S-measure gains when the synthetic pairs are added to a small training set.
-
SAM-Mamba: Mamba Guided SAM Architecture for Generalized Zero-Shot Polyp Segmentation
SAM-Mamba couples a Mamba-based prior with a frozen SAM encoder and adapter fine-tuning to achieve state-of-the-art polyp segmentation and cross-dataset zero-shot generalization.
-
ToonOut: Fine-tuned Background-Removal for Anime Characters
Fine-tuning BiRefNet on a small synthetic anime dataset lifts their test-set pixel accuracy from 95.3% to 99.5%, but the test set is curated from the same distribution.
-
HiddenObject: Modality-Agnostic Fusion for Multimodal Hidden Object Detection
A Mamba-based fusion network with a channel-aware decoder reports competitive or state-of-the-art results on RGB-thermal and RGB-depth hidden-object detection benchmarks.
-
Lightweight Multi-Scale Feature Extraction with Fully Connected LMF Layer for Salient Object Detection
LMFNet is a lightweight saliency detector built from depthwise separable dilated convolutions, but its results and implementation do not match the paper's central claims.
-
Dual Mutual Learning Network with Global-local Awareness for RGB-D Salient Object Detection
GL-DMNet combines RGB and depth features through position and channel mutual fusion modules and a transformer-infused decoder, reporting state-of-the-art averages on six RGB-D salient object detection benchmarks.
-
Structure-Aware Stylized Image Synthesis for Robust Medical Image Segmentation
Stylizing training images to the test-domain style with structure-preserving diffusion improves segmentation scores, but the protocol uses target-domain images for style transfer.
Discussion (0). Continue with ORCID to comment.