Pith. sign in

REVIEW 6 cited by

PMLB v1.0: An open source dataset collection for benchmarking machine learning methods

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2012.00058 v3 pith:SMIAWLPP submitted 2020-11-30 cs.LG cs.DB

classification cs.LGcs.DB
keywords pmlbdatasetslearningmachinemethodsbenchmarkcollectiondata
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Motivation: Novel machine learning and statistical modeling studies rely on standardized comparisons to existing methods using well-studied benchmark datasets. Few tools exist that provide rapid access to many of these datasets through a standardized, user-friendly interface that integrates well with popular data science workflows. Results: This release of PMLB provides the largest collection of diverse, public benchmark datasets for evaluating new machine learning and data science methods aggregated in one location. v1.0 introduces a number of critical improvements developed following discussions with the open-source community. Availability: PMLB is available at https://github.com/EpistasisLab/pmlb. Python and R interfaces for PMLB can be installed through the Python Package Index and Comprehensive R Archive Network, respectively.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 6 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Probabilistic Pretraining for Neural Regression

    cs.LG 2025-08 reject novelty 6.0 of 10

    Pretraining a permutation-invariant quantile network across 101 tabular datasets improves fine-tuned accuracy and calibration, but the headline claim of beating well-tuned tree ensembles is contradicted by the paper's...

  2. Are machine learning interpretations reliable? A stability study on global interpretations

    stat.ML 2025-05 conditional novelty 6.0 of 10

    Popular machine learning interpretation methods are frequently unstable under small data perturbations, and interpretation stability does not track prediction accuracy.

  3. Transformer Semantic Genetic Programming for Symbolic Regression

    cs.NE 2025-01 conditional novelty 6.0 of 10

    A transformer trained on synthetic function pairs with similar behavior can act as a semantic variation operator for genetic programming, improving symbolic regression accuracy and solution size.

  4. Synthesizing real-world distributions from high-dimensional Gaussian Noise with Fully Connected Neural Network

    cs.LG 2026-04 unverdicted novelty 5.0 of 10

    Fully connected neural network with randomized loss synthesizes real-world tabular data distributions from Gaussian noise faster than state-of-the-art deep generative models.

  5. What should an AI assessor optimise for?

    cs.LG 2025-02 conditional novelty 5.0 of 10

    Proxy loss functions can outperform the target loss when training AI assessors, with logistic loss and logarithmic score being the most promising proxies in the reported experiments.

  6. Constrained Hybrid Metaheuristic Algorithm for Probabilistic Neural Networks Learning

    cs.NE 2025-01 reject novelty 5.0 of 10

    A probe-then-fit portfolio of five metaheuristics attains the best rank in average test accuracy on 16 benchmarks for probabilistic neural networks, but the test set appears to be used as the training objective.

Pith tools