Pith. sign in

REVIEW 2 cited by

AutoElicit: Using Large Language Models for Expert Prior Elicitation in Predictive Modelling

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2411.17284 v5 pith:M2SEKCBA submitted 2024-11-26 cs.LG cs.CLstat.ML

classification cs.LGcs.CLstat.ML
keywords autoelicitpriorsmodelspredictivelanguagelearningmodelcomplexity
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Large language models (LLMs) acquire a breadth of information across various domains. However, their computational complexity, cost, and lack of transparency often hinder their direct application for predictive tasks where privacy and interpretability are paramount. In fields such as healthcare, biology, and finance, specialised and interpretable linear models still hold considerable value. In such domains, labelled data may be scarce or expensive to obtain. Well-specified prior distributions over model parameters can reduce the sample complexity of learning through Bayesian inference; however, eliciting expert priors can be time-consuming. We therefore introduce AutoElicit to extract knowledge from LLMs and construct priors for predictive models. We show these priors are informative and can be refined using natural language. We perform a careful study contrasting AutoElicit with in-context learning and demonstrate how to perform model selection between the two methods. We find that AutoElicit yields priors that can substantially reduce error over uninformative priors, using fewer labels, and consistently outperform in-context learning. We show that AutoElicit saves over 6 months of labelling effort when building a new predictive model for urinary tract infections from sensor recordings of people living with dementia.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Exploiting LLMs for Automatic Hypothesis Assessment via a Logit-Based Calibrated Prior

    cs.LG 2025-06 conditional novelty 5.0 of 10

    A logit-based method converts an LLM's numeric guesses into a calibrated prior over Pearson correlations and ranks expert-flagged hypotheses better than ranking by magnitude or by a fine-tuned RoBERTa classifier.

  2. Using Large Language Models to Suggest Informative Prior Distributions in Bayesian Statistics

    stat.ME 2025-06 conditional novelty 4.0 of 10

    LLMs suggested directionally correct but poorly calibrated Bayesian priors, with Claude's weak priors ranking best on KL divergence from the data.

Pith tools