REVIEW 2 cited by
Curriculum Learning by Transfer Learning: Theory and Experiments with Deep Networks
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Curriculum Learning by Transfer Learning: Theory and Experiments with Deep Networks
read the original abstract
We provide theoretical investigation of curriculum learning in the context of stochastic gradient descent when optimizing the convex linear regression loss. We prove that the rate of convergence of an ideal curriculum learning method is monotonically increasing with the difficulty of the examples. Moreover, among all equally difficult points, convergence is faster when using points which incur higher loss with respect to the current hypothesis. We then analyze curriculum learning in the context of training a CNN. We describe a method which infers the curriculum by way of transfer learning from another network, pre-trained on a different task. While this approach can only approximate the ideal curriculum, we observe empirically similar behavior to the one predicted by the theory, namely, a significant boost in convergence speed at the beginning of training. When the task is made more difficult, improvement in generalization performance is also observed. Finally, curriculum learning exhibits robustness against unfavorable conditions such as excessive regularization.
Forward citations
Cited by 2 Pith papers
-
LiteMatch: Lightweight Zero-Shot Stereo Matching via Cost Volume Stabilization
LiteMatch uses CVCE and HFE encoders plus CVC-Loss to enable lightweight zero-shot stereo matching with competitive EPE and D1 scores on Scene Flow, KITTI, Middlebury, ETH3D, and DrivingStereo using 3.36M-9.58M parameters.
-
Learning to Reason at the Frontier of Learnability
A curriculum sampling questions with high variance in success rate improves reinforcement learning performance for LLM reasoning tasks.
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.