Pith. sign in

REVIEW 2 cited by

Subgroup Robustness Grows On Trees: An Empirical Baseline Investigation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2211.12703 v2 pith:6OK6MLSH submitted 2022-11-23 cs.LG cs.CY

classification cs.LGcs.CY
keywords tree-basedmethodsmodelsrobustnesssubgroupempiricalrobustbaseline
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

Researchers have proposed many methods for fair and robust machine learning, but comprehensive empirical evaluation of their subgroup robustness is lacking. In this work, we address this gap in the context of tabular data, where sensitive subgroups are clearly-defined, real-world fairness problems abound, and prior works often do not compare to state-of-the-art tree-based models as baselines. We conduct an empirical comparison of several previously-proposed methods for fair and robust learning alongside state-of-the-art tree-based methods and other baselines. Via experiments with more than $340{,}000$ model configurations on eight datasets, we show that tree-based methods have strong subgroup robustness, even when compared to robustness- and fairness-enhancing methods. Moreover, the best tree-based models tend to show good performance over a range of metrics, while robust or group-fair models can show brittleness, with significant performance differences across different metrics for a fixed model. We also demonstrate that tree-based models show less sensitivity to hyperparameter configurations, and are less costly to train. Our work suggests that tree-based ensemble models make an effective baseline for tabular data, and are a sensible default when subgroup robustness is desired. For associated code and detailed results, see https://github.com/jpgard/subgroup-robustness-grows-on-trees .

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Learning-to-Context Slope: Evaluating In-Context Learning Effectiveness Beyond Performance Illusions

    cs.CL 2025-06 conditional novelty 5.0 of 10

    LCS, a slope between learning gain and contextual relevance, is proposed and shown to correlate with ICL performance gains, with a suggested threshold of 0.2 for effective in-context learning.

  2. Data Heterogeneity Modeling for Trustworthy Machine Learning

    cs.LG 2025-06 conditional novelty 3.0 of 10

    A survey that frames heterogeneity-aware machine learning as a paradigm spanning data collection, training, evaluation, and deployment, drawing mostly on the authors' prior results.

Pith tools