REVIEW 3 cited by
HyperbolicLR: Epoch insensitive learning rate scheduler
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This study proposes two novel learning rate schedulers -- Hyperbolic Learning Rate Scheduler (HyperbolicLR) and Exponential Hyperbolic Learning Rate Scheduler (ExpHyperbolicLR) -- to address the epoch sensitivity problem that often causes inconsistent learning curves in conventional methods. By leveraging the asymptotic behavior of hyperbolic curves, the proposed schedulers maintain more stable learning curves across varying epoch settings. Specifically, HyperbolicLR applies this property directly in the epoch-learning rate space, while ExpHyperbolicLR extends it to an exponential space. We first determine optimal hyperparameters for each scheduler on a small number of epochs, fix these hyperparameters, and then evaluate performance as the number of epochs increases. Experimental results on various deep learning tasks (e.g., image classification, time series forecasting, and operator learning) demonstrate that both HyperbolicLR and ExpHyperbolicLR achieve more consistent performance improvements than conventional schedulers as training duration grows. These findings suggest that our hyperbolic-based schedulers offer a more robust and efficient approach to deep network optimization, particularly in scenarios constrained by computational resources or time.
Forward citations
Cited by 3 Pith papers
-
FedA2L: Adaptive layer-wise learning rate adjustment in decentralized federated learning
FedA2L adapts per-layer learning rates from local weight-divergence and aggregation-stability signals, accelerating convergence in decentralized federated learning without extra communication.
-
CIKAN: Constraint Informed Kolmogorov-Arnold Networks for Autonomous Spacecraft Rendezvous using Time Shift Governor
CIKAN, a KAN-based constraint-informed network, approximates the Time Shift Governor for spacecraft rendezvous and, in simulation, enforces constraints while reducing average computation time and fuel use relative to ...
-
Learning Hamiltonian Dynamics with Bayesian Data Assimilation
An autoregressive Hamiltonian neural network coupled with an unscented Kalman filter improves long-term trajectory prediction and uncertainty quantification for unknown Hamiltonian systems.
Discussion (0). Continue with ORCID to comment.