REVIEW 9 cited by
FCNs in the Wild: Pixel-level Adversarial and Constraint-based Adaptation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Fully convolutional models for dense prediction have proven successful for a wide range of visual tasks. Such models perform well in a supervised setting, but performance can be surprisingly poor under domain shifts that appear mild to a human observer. For example, training on one city and testing on another in a different geographic region and/or weather condition may result in significantly degraded performance due to pixel-level distribution shift. In this paper, we introduce the first domain adaptive semantic segmentation method, proposing an unsupervised adversarial approach to pixel prediction problems. Our method consists of both global and category specific adaptation techniques. Global domain alignment is performed using a novel semantic segmentation network with fully convolutional domain adversarial learning. This initially adapted space then enables category specific adaptation through a generalization of constrained weak learning, with explicit transfer of the spatial layout from the source to the target domains. Our approach outperforms baselines across different settings on multiple large-scale datasets, including adapting across various real city environments, different synthetic sub-domains, from simulated to real environments, and on a novel large-scale dash-cam dataset.
Forward citations
Cited by 9 Pith papers
-
Blending-target Domain Adaptation by Adversarial Meta-Adaptation Networks
AMEAN applies adversarial meta-learning to discover implicit meta-sub-target clusters in blended target data, reducing intra-target category misalignment and outperforming standard DA methods on three BTDA benchmarks.
-
PairedGTA: Generating Driving Datasets for Controlled Photometric Shift Analysis
PairedGTA uses a game engine to produce pixel-aligned driving images under controlled photometric variations for isolated evaluation of environmental effects on perception models.
-
Make me an Expert: Distilling from Generalist Black-Box Models into Specialized Models for Semantic Segmentation
ATGC selects the best input scale for a black-box open-vocabulary segmentation API, using DINOv2 attention entropy, improving one-hot-label distillation on Cityscapes and ACDC.
-
Boundary and Entropy-driven Adversarial Learning for Fundus Image Segmentation
BEAL improves cross-domain optic disc and cup segmentation by adversarially aligning boundary predictions and entropy maps from source to target domains.
-
Interact3D: Compositional 3D Generation of Interactive Objects
Interact3D composes multi-object 3D scenes from one image via asset generation, registration plus SDF collision penalties, and VLM-driven closed-loop image edits.
-
$\varphi$-Adapt: A Physics-Informed Adaptation Learning Approach to 2D Quantum Material Discovery
A physics-informed domain adaptation approach, trained on 600,000 synthesized flake images, is claimed to set state-of-the-art results for detection, layer classification, and thickness estimation on real 2D material ...
-
Real-Time Per-Garment Virtual Try-On with Temporal Consistency for Loose-Fitting Garments
A per-garment virtual try-on method for loose-fitting garments uses a garment-invariant pose representation and a recurrent ConvLSTM synthesis network to achieve temporally smoother try-on video at about 10 fps.
-
TITAN: Query-Token based Domain Adaptive Adversarial Learning
TITAN claims large source-free domain adaptation gains using variance-based target splitting and query-token adversarial alignment, but internal contradictions and test-set leakage invalidate the reported results.
-
How much real data do we actually need: Analyzing object detection performance using synthetic and real data
Synthetic data can partially substitute for real data in object detection training, with performance tied to domain similarity and the volume of real data included.
Discussion (0). Sign in to comment.