REVIEW 4 cited by
A survey and taxonomy of loss functions in machine learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Most state-of-the-art machine learning techniques revolve around the optimisation of loss functions. Defining appropriate loss functions is therefore critical to successfully solving problems in this field. In this survey, we present a comprehensive overview of the most widely used loss functions across key applications, including regression, classification, generative modeling, ranking, and energy-based modeling. We introduce 43 distinct loss functions, structured within an intuitive taxonomy that clarifies their theoretical foundations, properties, and optimal application contexts. This survey is intended as a resource for undergraduate, graduate, and Ph.D. students, as well as researchers seeking a deeper understanding of loss functions.
Forward citations
Cited by 4 Pith papers
-
A Unified Framework for In-Context Learning with Causal and Masked Language Models
Masked and causal pretraining yield same-order k-shot excess-risk bounds under Wasserstein regularity, and a Masked Pair Encoder matches GPT-2-style ICL on synthetic function classes.
-
Reconstruction of Primordial Power Spectrum from Gravitational Waves of High-Redshift Black Hole Binaries
Gradient-descent inversion of redshifted BBH mass distributions recovers the PBH mass function and, via regularized Press-Schechter, a candidate O(10^{-2}) bump in the small-scale primordial power spectrum.
-
Neuro-Argumentative Learning with Case-Based Reasoning
Gradual AA-CBR learns a case-based argumentation debate end-to-end and matches a single-layer neural network on four datasets while supporting multi-class and continuous features.
-
Bulk-boundary decomposition of neural networks
The paper reframes SGD training of deep networks as a local Lagrangian with data confined to the boundaries, but the advertised energy continuity equation is absent from the body.
Discussion (0). Continue with ORCID to comment.