Pith. sign in

REVIEW

A New Optimal Stepsize For Approximate Dynamic Programming

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1407.2676 v2 pith:QQL2CDJW submitted 2014-07-10 math.OC cs.AIcs.LGcs.SYeess.SYstat.ML

classification math.OCcs.AIcs.LGcs.SYeess.SYstat.ML
keywords stepsizeruleapplicationsapproximatedynamicmanyprogrammingresults
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Approximate dynamic programming (ADP) has proven itself in a wide range of applications spanning large-scale transportation problems, health care, revenue management, and energy systems. The design of effective ADP algorithms has many dimensions, but one crucial factor is the stepsize rule used to update a value function approximation. Many operations research applications are computationally intensive, and it is important to obtain good results quickly. Furthermore, the most popular stepsize formulas use tunable parameters and can produce very poor results if tuned improperly. We derive a new stepsize rule that optimizes the prediction error in order to improve the short-term performance of an ADP algorithm. With only one, relatively insensitive tunable parameter, the new rule adapts to the level of noise in the problem and produces faster convergence in numerical experiments.

Discussion (0). Continue with ORCID to comment.

Pith tools