Pith. sign in

Artificial Constraints and Lipschitz Hints for Unconstrained Online Learning

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

We provide algorithms that guarantee regret $R_T(u)\le \tilde O(G\|u\|^3 + G(\|u\|+1)\sqrt{T})$ or $R_T(u)\le \tilde O(G\|u\|^3T^{1/3} + GT^{1/3}+ G\|u\|\sqrt{T})$ for online convex optimization with $G$-Lipschitz losses for any comparison point $u$ without prior knowledge of either $G$ or $\|u\|$. Previous algorithms dispense with the $O(\|u\|^3)$ term at the expense of knowledge of one or both of these parameters, while a lower bound shows that some additional penalty term over $G\|u\|\sqrt{T}$ is necessary. Previous penalties were exponential while our bounds are polynomial in all quantities. Further, given a known bound $\|u\|\le D$, our same techniques allow us to design algorithms that adapt optimally to the unknown value of $\|u\|$ without requiring knowledge of $G$.

citation-role summary

background 1

citation-polarity summary

fields

cs.RO 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

background 1

representative citing papers

Online Aggregation of Trajectory Predictors

cs.RO · 2025-02-11 · conditional · novelty 5.0

An online learning rule, based on SQUINT, mixes multiple trajectory predictors and tracks the best expert under distribution shift.

citing papers explorer

Showing 1 of 1 citing paper.

  • Online Aggregation of Trajectory Predictors cs.RO · 2025-02-11 · conditional · none · ref 16 · internal anchor

    An online learning rule, based on SQUINT, mixes multiple trajectory predictors and tracks the best expert under distribution shift.