Pith. sign in

REVIEW 1 cited by

Variational Adaptive-Newton Method for Explorative Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1711.05560 v1 pith:KYD5COED submitted 2017-11-15 stat.ML cs.LG

classification stat.MLcs.LG
keywords learningmethodmethodsoptimizationvariationalactivecontinuousexplorative-learning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present the Variational Adaptive Newton (VAN) method which is a black-box optimization method especially suitable for explorative-learning tasks such as active learning and reinforcement learning. Similar to Bayesian methods, VAN estimates a distribution that can be used for exploration, but requires computations that are similar to continuous optimization methods. Our theoretical contribution reveals that VAN is a second-order method that unifies existing methods in distinct fields of continuous optimization, variational inference, and evolution strategies. Our experimental results show that VAN performs well on a wide-variety of learning tasks. This work presents a general-purpose explorative-learning method that has the potential to improve learning in areas such as active learning and reinforcement learning.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Optimization Guarantees for Square-Root Natural-Gradient Variational Inference

    cs.LG 2025-07 conditional novelty 7.0 of 10

    For strongly concave log-likelihoods, square-root (Cholesky) parametrization of Gaussian variational inference yields exponential convergence guarantees for both the natural-gradient flow and a discrete-time natural-g...

Pith tools