Pith. sign in

REVIEW 1 cited by

Priors for symbolic regression

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2304.06333 v2 pith:2PDMPEEP submitted 2023-04-13 cs.LG astro-ph.COastro-ph.IM

classification cs.LGastro-ph.COastro-ph.IM
keywords priorpriorssymbolicbayesiandevelopfunctionsmethodsmodel
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

When choosing between competing symbolic models for a data set, a human will naturally prefer the "simpler" expression or the one which more closely resembles equations previously seen in a similar context. This suggests a non-uniform prior on functions, which is, however, rarely considered within a symbolic regression (SR) framework. In this paper we develop methods to incorporate detailed prior information on both functions and their parameters into SR. Our prior on the structure of a function is based on a $n$-gram language model, which is sensitive to the arrangement of operators relative to one another in addition to the frequency of occurrence of each operator. We also develop a formalism based on the Fractional Bayes Factor to treat numerical parameter priors in such a way that models may be fairly compared though the Bayesian evidence, and explicitly compare Bayesian, Minimum Description Length and heuristic methods for model selection. We demonstrate the performance of our priors relative to literature standards on benchmarks and a real-world dataset from the field of cosmology.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. (Exhaustive) Symbolic Regression and model selection by minimum description length

    astro-ph.IM 2025-07 conditional novelty 3.0 of 10

    Exhaustive search over simple functions ranked by description length beats the Friedmann equation, MOND, and common inflaton potentials on astrophysical datasets.

Pith tools