REVIEW 2 cited by
Towards Stable Machine Learning Model Retraining via Slowly Varying Sequences
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We consider the problem of retraining machine learning (ML) models when new batches of data become available. Existing approaches greedily optimize for predictive power independently at each batch, without considering the stability of the model's structure or analytical insights across retraining iterations. We propose a model-agnostic framework for finding sequences of models that are stable across retraining iterations. We develop a mixed-integer optimization formulation that is guaranteed to recover Pareto optimal models (in terms of the predictive power-stability trade-off) with good generalization properties, as well as an efficient polynomial-time algorithm that performs well in practice. We focus on retaining consistent analytical insights-which is important to model interpretability, ease of implementation, and fostering trust with users-by using custom-defined distance metrics that can be directly incorporated into the optimization problem. We evaluate our framework across models (regression, decision trees, boosted trees, and neural networks) and application domains (healthcare, vision, and language), including deployment in a production pipeline at a major US hospital. We find that, on average, a 2% reduction in predictive power leads to a 30% improvement in stability.
Forward citations
Cited by 2 Pith papers
-
A Six-Dimensional Taxonomy of Post-Training Adaptation Techniques with Applications in AI Governance
A new taxonomy characterizes 48 post-training AI adaptation techniques on six axes and maps them to regulatory documentation requirements.
-
The Challenger: When Do New Data Sources Justify Switching Machine Learning Models?
Under a power-law learning curve, the optimal time to switch to a challenger model grows as T^{1/(1+alpha)}, and a look-ahead sequential algorithm empirically approaches an oracle with O(T^{2/3} sqrt(log T)) regret.
Discussion (0). Continue with ORCID to comment.