Pith. sign in

REVIEW

Policy Gradients for Optimal Parallel Tempering MCMC

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2409.01574 v2 pith:36UQLVIU submitted 2024-09-03 stat.CO cs.LGstat.ML

classification stat.COcs.LGstat.ML
keywords paralleltemperaturestemperingchaindistributionspolicyselectiontraditional
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Parallel tempering is a meta-algorithm for Markov Chain Monte Carlo that uses multiple chains to sample from tempered versions of the target distribution, enhancing mixing in multi-modal distributions that are challenging for traditional methods. The effectiveness of parallel tempering is heavily influenced by the selection of chain temperatures. Here, we present an adaptive temperature selection algorithm that dynamically adjusts temperatures during sampling using a policy gradient approach. Experiments demonstrate that our method can achieve lower integrated autocorrelation times compared to traditional geometrically spaced temperatures and uniform acceptance rate schemes on benchmark distributions.

Discussion (0). Continue with ORCID to comment.

Pith tools