REVIEW 10 cited by
Generalizing to unseen domains via distribution matching
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Supervised learning results typically rely on assumptions of i.i.d. data. Unfortunately, those assumptions are commonly violated in practice. In this work, we tackle such problem by focusing on domain generalization: a formalization where the data generating process at test time may yield samples from never-before-seen domains (distributions). Our work relies on the following lemma: by minimizing a notion of discrepancy between all pairs from a set of given domains, we also minimize the discrepancy between any pairs of mixtures of domains. Using this result, we derive a generalization bound for our setting. We then show that low risk over unseen domains can be achieved by representing the data in a space where (i) the training distributions are indistinguishable, and (ii) relevant information for the task at hand is preserved. Minimizing the terms in our bound yields an adversarial formulation which estimates and minimizes pairwise discrepancies. We validate our proposed strategy on standard domain generalization benchmarks, outperforming a number of recently introduced methods. Notably, we tackle a real-world application where the underlying data corresponds to multi-channel electroencephalography time series from different subjects, each considered as a distinct domain.
Forward citations
Cited by 10 Pith papers
-
ACE and Diverse Generalization via Selective Disagreement
ACE learns an ensemble of classifiers that agree on labeled data but confidently and selectively disagree on target-distribution data, recovering diverse human-interpretable concepts under complete spurious correlation.
-
Prototypical Progressive Alignment and Reweighting for Generalizable Semantic Segmentation
PPAR aligns segmentation features with CLIP text prototypes from easy to hard and reweights source pixels to improve generalization to unseen domains, achieving top average mIoU on several urban scene benchmarks.
-
A General Method for Detecting Information Generated by Large Language Models
A new LLM-text detector built from twin memory networks and domain-generalization losses outperforms prior detectors on held-out LLMs and domains in the paper's benchmark.
-
Multi-Domain Graph Foundation Models: Robust Knowledge Transfer via Topology Alignment
MDGFM aligns graph topologies across domains with learnable tokens and structure refinement, improving few-shot transfer to unseen graphs.
-
Towards Understanding Extrapolation: a Causal Lens
Under a minimal-change latent model, the invariant part of a representation is identifiable from one off-support target sample, under bounded dense shifts or arbitrary sparse shifts.
-
Open-Set Heterogeneous Domain Adaptation: Theoretical Analysis and Algorithm
RL-OSHeDA, a two-stage representation learning method with pseudo-labeling, outperforms existing domain adaptation baselines on 56 open-set heterogeneous domain adaptation tasks.
-
Cross-Receiver Generalization for RF Fingerprint Identification via Feature Disentanglement and Adversarial Training
DRIFT disentangles transmitter and receiver features and uses adversarial alignment plus receiver-clustering regularization to keep RF fingerprint identification accurate on unseen receivers.
-
Domain Generalization via Pareto Optimal Gradient Matching
POGM combines pairwise gradient inner product maximization with a ball constraint around the ERM gradient, using a meta-learning update that avoids second-order derivatives and achieves competitive DomainBed accuracy.
-
Weight Averaging for Out-of-Distribution Generalization and Few-Shot Domain Adaptation
Gradient-similarity-regularized weight averaging and WA+SAM fine-tuning are tested on OOD and few-shot domain adaptation benchmarks, with mixed results that do not support the claimed improvements.
-
Epistemic Artificial Intelligence is Essential for Machine Learning Models to Truly 'Know When They Do Not Know'
Machine learning should use second-order uncertainty measures, such as credal sets and random sets, so models can explicitly represent ignorance and avoid overconfident predictions on unfamiliar data.
Discussion (0). Continue with ORCID to comment.