REVIEW 18 cited by
Domain Generalization with MixStyle
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Though convolutional neural networks (CNNs) have demonstrated remarkable ability in learning discriminative features, they often generalize poorly to unseen domains. Domain generalization aims to address this problem by learning from a set of source domains a model that is generalizable to any unseen domain. In this paper, a novel approach is proposed based on probabilistically mixing instance-level feature statistics of training samples across source domains. Our method, termed MixStyle, is motivated by the observation that visual domain is closely related to image style (e.g., photo vs.~sketch images). Such style information is captured by the bottom layers of a CNN where our proposed style-mixing takes place. Mixing styles of training instances results in novel domains being synthesized implicitly, which increase the domain diversity of the source domains, and hence the generalizability of the trained model. MixStyle fits into mini-batch training perfectly and is extremely easy to implement. The effectiveness of MixStyle is demonstrated on a wide range of tasks including category classification, instance retrieval and reinforcement learning.
Forward citations
Cited by 18 Pith papers
-
Quantization Meets OOD: Generalizable Quantization-aware Training from a Flatness Perspective
Quantization-aware training degrades out-of-distribution accuracy, and a flatness-aware method with gradient-disorder freezing, FQAT, partially recovers it.
-
Vision Transformer Neural Architecture Search for Out-of-Distribution Generalization: Benchmark and Insights
A 3000-architecture benchmark shows ViT OoD accuracy varies widely with architecture and that embedding dimension is the strongest structural correlate of OoD robustness.
-
MoEcho: Exploiting Side-Channel Attacks to Compromise User Privacy in Mixture-of-Experts LLMs
MoEcho claims to compromise user privacy in MoE LLMs and VLMs via four CPU and GPU side channels, but the provided manuscript body contains no supporting content.
-
CorrMoE: Mixture of Experts with De-stylization Learning for Cross-Scene and Cross-Domain Correspondence Pruning
CorrMoE combines Progressive Mixstyle de-stylization with a Bi-Fusion Mixture-of-Experts module to improve two-view correspondence pruning in cross-domain and cross-scene settings.
-
CTA: Cross-Task Alignment for Better Test Time Training
CTA aligns a SimCLR-trained encoder to a frozen supervised encoder, then adapts only the self-supervised encoder at test time, improving corrupted-image classification over prior test-time training methods.
-
DAM: Domain-Aware Module for Multi-Domain Dataset Condensation
A training-time module with learnable spatial masks and frequency-based pseudo-domain labels improves dataset condensation on multi-domain data without increasing images per class.
-
Seeking Consistent Flat Minima for Better Domain Generalization via Refining Loss Landscapes
A self-feedback training framework that refines loss landscapes with dynamically generated soft labels finds more consistent flat minima and improves domain generalization accuracy across five benchmarks.
-
PIER: Physics-Informed Environmental Retrieval for Time-Series Modeling
PIER augments embedding-based retrieval for lake modeling with a physics-aware stream scored by local verifiers, improving water temperature and dissolved oxygen prediction across 356 lakes.
-
Towards Realistic Hand-Object Interaction with Gravity-Field Based Diffusion Bridge
A gravity-field and diffusion-based optimization refines hand-object contacts to reduce interpenetration and gaps while LLM text prompts guide contact regions.
-
Learning Semantic Directions for Feature Augmentation in Domain-Generalized Medical Segmentation
A feature-augmentation framework with a learned channel selector and covariance-guided intensity sampler that reports the best average domain-generalized segmentation on Prostate and Fundus benchmarks.
-
Positive Style Accumulation: A Style Screening and Continuous Utilization Framework for Federated DG-ReID
SSCU improves federated domain-generalizable person re-identification by screening styles with round-over-round Rank-1 gains and continuously training on the memorized positive styles.
-
Boosting Domain Generalized and Adaptive Detection with Diffusion Models: Fitness, Generalization, and Transferability
A single-step diffusion feature extractor with an object-masked auxiliary branch and consistency loss improves domain-generalized and adaptive detection accuracy and speed.
-
Adaptive Knowledge Distillation using a Device-Aware Teacher for Low-Complexity Acoustic Scene Classification
A low-complexity scene classifier trained by two-teacher knowledge distillation and device-specific fine-tuning reaches 57.93% accuracy on the DCASE 2025 development set.
-
Mix, Align, Distil: Reliable Cross-Domain Atypical Mitosis Classification
A DenseNet-121 trained with MixStyle, CBAM-based feature alignment, and EMA-teacher distillation achieves 0.8762 balanced accuracy on the MIDOG 2025 Task 2 atypical mitosis classification leaderboard.
-
Fully Automated SAM for Single-source Domain Generalization in Medical Image Segmentation
FA-SAM automates SAM-based medical segmentation across domains by generating prompt boxes with an uncertainty-enhanced network and fusing image and prompt embeddings.
-
Fourier Asymmetric Attention on Domain Generalization for Pan-Cancer Drug Response Prediction
FourierDrug uses bulk cell-line expression with adversarial domain generalization and a Fourier asymmetric attention constraint to predict drug response in unseen cancer types, single cells, and patients.
-
Spatially-Delineated Domain-Adapted AI Classification: An Application for Oncology Data
A multi-task self-supervised framework with spatial mix-up masking and contrastive predictive coding improves unsupervised domain adaptation for classifying multi-type point maps from cancer tissue regions.
-
Improving Acoustic Scene Classification in Low-Resource Conditions
DS-FlexiNet achieves 58.25% accuracy after int8 quantization on TAU22 Task 1A with 30.69K parameters and 8.27M MACs, using residual normalization, ADIR augmentation, and 12-teacher knowledge distillation.
Discussion (0). Continue with ORCID to comment.