Pith. sign in

REVIEW 3 cited by

LightAutoML: AutoML Solution for a Large Financial Services Ecosystem

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2109.01528 v2 pith:EPWBRYB3 submitted 2021-09-03 cs.LG stat.ML

classification cs.LGstat.ML
keywords automlecosystemsystemdatafinanciallargelightautomlscientists
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present an AutoML system called LightAutoML developed for a large European financial services company and its ecosystem satisfying the set of idiosyncratic requirements that this ecosystem has for AutoML solutions. Our framework was piloted and deployed in numerous applications and performed at the level of the experienced data scientists while building high-quality ML models significantly faster than these data scientists. We also compare the performance of our system with various general-purpose open source AutoML solutions and show that it performs better for most of the ecosystem and OpenML problems. We also present the lessons that we learned while developing the AutoML system and moving it into production.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. ML2B: Benchmarking LLMs on Cross-Lingual ML Pipeline Generation

    cs.CL 2025-09 conditional novelty 6.0 of 10

    A new open-source benchmark evaluates LLM-generated end-to-end ML pipelines from Kaggle competition descriptions translated into 13 languages, with 6 private tasks to limit data leakage.

  2. Imbalanced Regression Pipeline Recommendation

    cs.LG 2025-07 conditional novelty 6.0 of 10

    Meta-IR trains meta-classifiers on 218 datasets to recommend a regression model and a resampling strategy from dataset meta-features, with a chained variant that slightly improves F1-scoreR but not SERA.

  3. Interpretable by Design: MH-AutoML for Transparent and Efficient Android Malware Detection without Compromising Performance

    cs.CR 2025-06 conditional novelty 5.0 of 10

    MH-AutoML is a domain-specific AutoML framework for Android malware detection that combines automated modeling with built-in interpretability, and its evaluation shows competitive recall and higher transparency scores...

Pith tools