Pith. sign in

REVIEW 16 cited by

Bayesian Workflow

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2011.01808 v1 pith:QXIVT3X4 submitted 2020-11-03 stat.ME

Bayesian Workflow

classification stat.ME
keywords modelbayesianworkflowmanymodelsanalysisaspectsdata
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
Share X Bluesky LinkedIn Reddit HN
read the original abstract

The Bayesian approach to data analysis provides a powerful way to handle uncertainty in all observations, model parameters, and model structure using probability theory. Probabilistic programming languages make it easier to specify and fit Bayesian models, but this still leaves us with many options regarding constructing, evaluating, and using these models, along with many remaining challenges in computation. Using Bayesian inference to solve real-world problems requires not only statistical skills, subject matter knowledge, and programming, but also awareness of the decisions made in the process of data analysis. All of these aspects can be understood as part of a tangled workflow of applied Bayesian statistics. Beyond inference, the workflow also includes iterative model building, model checking, validation and troubleshooting of computational problems, model understanding, and model comparison. We review all these aspects of workflow in the context of several examples, keeping in mind that in practice we will be fitting many models for any given problem, even if only a subset of them will ultimately be relevant for our conclusions.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 16 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Calibration, Not Compilation: Detecting and Repairing Misspecified Probabilistic Programs Written by Language Models

    cs.LG 2026-06 unverdicted novelty 7.0

    Bayesian workflow diagnostics outperform unit tests for detecting and repairing statistically misspecified LLM-generated probabilistic programs across benchmarks and real generation tasks.

  2. Beyond Empirical Bayes: A Hierarchical Bayesian Approach to Crash Rate Estimation with Missing Traffic Volume

    stat.AP 2026-05 unverdicted novelty 7.0

    A hierarchical Bayesian model jointly imputes missing ADT and estimates segment crash rates with flexible exposure exponents, outperforming Empirical Bayes by PSIS-LOO Δelpd of 9,394 on Ohio data.

  3. Toward Joint Prediction of a Longitudinal Marker and a Terminal Event: A bivariate discrete-time framework

    stat.ME 2026-07 conditional novelty 6.0

    A flexible Bayesian bivariate discrete-time model jointly predicts partly-conditional longitudinal trajectories and terminal-event risk without immortal-cohort extrapolation.

  4. To select or not to select: predictively consistent priors instead of model selection

    stat.ME 2026-06 unverdicted novelty 6.0

    Predictively consistent priors let complex Bayesian models match or beat the out-of-sample performance of selected simpler models across linear, logistic, and nonlinear examples without explicit selection.

  5. A nutritionally informed model for Bayesian variable selection with metabolite response variables

    stat.ME 2025-09 conditional novelty 6.0

    A new Bayesian variable selection method for metabolite data, using a skew-normal censored mixture with a Markov random field prior, detects diet-metabolite associations in two cohorts.

  6. RefineStat: Efficient Exploration for Probabilistic Program Synthesis

    cs.LG 2025-09 unverdicted novelty 6.0

    RefineStat improves small language model performance on probabilistic program synthesis by adding semantic constraint enforcement and diagnostic-aware refinement, producing syntactically and statistically reliable cod...

  7. Compositional amortized inference for large-scale hierarchical Bayesian models

    q-bio.QM 2025-05 unverdicted novelty 6.0

    A new error-damping estimator for compositional score matching enables stable amortized inference on hierarchical Bayesian models with over 750,000 parameters using fewer than one full model simulation on large problems.

  8. Toward Joint Prediction of a Longitudinal Marker and a Terminal Event: A bivariate discrete-time framework

    stat.ME 2026-07 conditional novelty 5.0

    A Bayesian discrete-time joint model predicts a patient's partly-conditional longitudinal marker trajectory together with the time to a terminal event, illustrated on ICD patients' quality of life and mortality.

  9. Bridging electrode preparation and electrocatalyst performance with physics-based causal AI

    cond-mat.mtrl-sci 2026-06 unverdicted novelty 5.0

    Physics-based structural causal models are used on n<10 multi-modal datasets to quantitatively disentangle support-to-catalyst ratio and loading effects on manganese-antimony oxide ORR performance in alkaline RDE tests.

  10. Restricted Multivariate Spatial Modeling

    stat.ME 2026-06 unverdicted novelty 5.0

    Develops a restricted MCAR model via reparameterization to measure and control informativeness in multivariate spatial modeling of health events across subgroups.

  11. How Requirements Quality Makes (or Breaks) Traceability Link Recovery

    cs.SE 2026-06 unverdicted novelty 5.0

    Empirical analysis of 189 annotated use cases shows that some requirements quality defects reduce TLR performance while others improve it, with effects varying by approach type.

  12. Quantifying Evidential Rigor in Meta-Analytic Corpora: A Simulation-Characterized, Bias-Robust Bayesian Workflow with a Nutrition Case Study

    stat.ME 2026-05 unverdicted novelty 5.0

    Introduces a corpus-scale Bayesian evidential-audit workflow that defines rigor as combined Bayes-factor evidence for effect or no-effect plus absence of explicit bias components, validated via simulations and applied...

  13. Bayesian Inference of Discretization Error Means in ODEs via Ensemble Kalman Filtering

    math.NA 2026-07 conditional novelty 4.0

    A Bayesian state-space model with an Ensemble Kalman Filter infers the mean of ODE discretization errors from noisy observations, using a step-size-dependent Markov prior whose convergence is proven.

  14. Comparative study of Bayesian and Frequentist methods for epidemic forecasting: Insights from simulated and historical data

    q-bio.QM 2025-09 reject novelty 4.0

    Neither Bayesian nor frequentist fitting is uniformly better for epidemic forecasts; performance depends on phase and data, though the paper's own results undercut its phase-specific claims.

  15. A Tutorial on Bayesian Analysis of Linear Shock Compression Data

    stat.AP 2026-03 conditional novelty 3.0

    Bayesian linear regression yields an analytic t-distribution posterior for C0 and S, which can be sampled and pushed through Rankine-Hugoniot equations to obtain pressure-volume Hugoniot credible intervals.

  16. An Introduction to Bayesian and Frequentist Simulation-Based Inference with Machine Learning

    cs.LG 2026-07 accept novelty 2.0

    A structured introduction to ML-based simulation-based inference, contrasting Bayesian and frequentist frameworks and covering parameter inference, unfolding, and validation.