REVIEW 3 cited by
A survey on domain adaptation theory: learning bounds and theoretical guarantees
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
All famous machine learning algorithms that comprise both supervised and semi-supervised learning work well only under a common assumption: the training and test data follow the same distribution. When the distribution changes, most statistical models must be reconstructed from newly collected data, which for some applications can be costly or impossible to obtain. Therefore, it has become necessary to develop approaches that reduce the need and the effort to obtain new labeled samples by exploiting data that are available in related areas, and using these further across similar fields. This has given rise to a new machine learning framework known as transfer learning: a learning setting inspired by the capability of a human being to extrapolate knowledge across tasks to learn more efficiently. Despite a large amount of different transfer learning scenarios, the main objective of this survey is to provide an overview of the state-of-the-art theoretical results in a specific, and arguably the most popular, sub-field of transfer learning, called domain adaptation. In this sub-field, the data distribution is assumed to change across the training and the test data, while the learning task remains the same. We provide a first up-to-date description of existing results related to domain adaptation problem that cover learning bounds based on different statistical learning frameworks.
Forward citations
Cited by 3 Pith papers
-
A Unified Analysis of Generalization and Sample Complexity for Semi-Supervised Domain Adaptation
The sample complexity of MMD and adversarial domain-adaptive networks is upper-bounded by O(d^2 L^2 / eps^2), and the target-loss weight should scale as O(sqrt(M_t)).
-
When few labeled target data suffice: a theory of semi-supervised domain adaptation via fine-tuning from multiple adaptive starts
Under linear anticausal causal models, fine-tuning from UDA starts achieves target-label sample complexity proportional to the intervention dimension, not the ambient dimension.
-
DataPrep-Bench: Benchmarking LLMs as Training Data Preparators
DataPrep-Bench jointly benchmarks data construction and data-quality evaluation for LLMs across six domains with downstream fine-tuning performance as ground truth.
Discussion (0). Continue with ORCID to comment.