A flaky-test classifier scores 8% higher on augmented copies of its training data than on independent test cases, but the comparison conflates augmentation with train-test overlap.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SE 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
Assessing Data Augmentation-Induced Bias in Training and Testing of Machine Learning Models
A flaky-test classifier scores 8% higher on augmented copies of its training data than on independent test cases, but the comparison conflates augmentation with train-test overlap.