Pith. sign in

REVIEW 1 cited by

Artificial Constraints and Lipschitz Hints for Unconstrained Online Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1902.09013 v1 pith:WPLELGVP submitted 2019-02-24 stat.ML cs.LGmath.OC

classification stat.MLcs.LGmath.OC
keywords algorithmsknowledgesqrtboundlipschitzonlinepreviousterm
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

We provide algorithms that guarantee regret $R_T(u)\le \tilde O(G\|u\|^3 + G(\|u\|+1)\sqrt{T})$ or $R_T(u)\le \tilde O(G\|u\|^3T^{1/3} + GT^{1/3}+ G\|u\|\sqrt{T})$ for online convex optimization with $G$-Lipschitz losses for any comparison point $u$ without prior knowledge of either $G$ or $\|u\|$. Previous algorithms dispense with the $O(\|u\|^3)$ term at the expense of knowledge of one or both of these parameters, while a lower bound shows that some additional penalty term over $G\|u\|\sqrt{T}$ is necessary. Previous penalties were exponential while our bounds are polynomial in all quantities. Further, given a known bound $\|u\|\le D$, our same techniques allow us to design algorithms that adapt optimally to the unknown value of $\|u\|$ without requiring knowledge of $G$.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Online Aggregation of Trajectory Predictors

    cs.RO 2025-02 conditional novelty 5.0 of 10

    An online learning rule, based on SQUINT, mixes multiple trajectory predictors and tracks the best expert under distribution shift.

Pith tools