REVIEW 3 cited by
Imbalance in Regression Datasets
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
For classification, the problem of class imbalance is well known and has been extensively studied. In this paper, we argue that imbalance in regression is an equally important problem which has so far been overlooked: Due to under- and over-representations in a data set's target distribution, regressors are prone to degenerate to naive models, systematically neglecting uncommon training data and over-representing targets seen often during training. We analyse this problem theoretically and use resulting insights to develop a first definition of imbalance in regression, which we show to be a generalisation of the commonly employed imbalance measure in classification. With this, we hope to turn the spotlight on the overlooked problem of imbalance in regression and to provide common ground for future research.
Forward citations
Cited by 3 Pith papers
-
Instance Hardness-Based Relevance for Imbalanced Regression
An instance-hardness relevance function that labels examples by prediction difficulty improves oversampling for imbalanced regression, with modest empirical gains.
-
Polar coordinate transformations for machine learning based dark matter subhalo detection in strong gravitational lenses
Polar-transformed strong-lensing images raise CNN subhalo detection fractions by ~15% relative to Cartesian inputs for 10^9–10^9.5 solar-mass subhalos on simulated HST data.
-
Model-agnostic Mitigation Strategies of Data Imbalance for Regression
The paper proposes two relevance functions and two sampling methods for imbalanced regression, and reports that crbSMOGN with density-ratio relevance improves rare-sample prediction for neural networks, while an ensem...
Discussion (0). Continue with ORCID to comment.