REVIEW 6 cited by
Learning subgaussian classes : Upper and minimax bounds
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
We obtain sharp oracle inequalities for the empirical risk minimization procedure in the regression model under the assumption that the target Y and the model F are subgaussian. The bound we obtain is sharp in the minimax sense if F is convex. Moreover, under mild assumptions on F, the error rate of ERM remains optimal even if the procedure is allowed to perform with constant probability. A part of our analysis is a new proof of minimax results for the gaussian regression model.
Forward citations
Cited by 6 Pith papers
-
Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models
For diffusion model fine-tuning, RL value estimation reduces to a variational inequality whose solution satisfies a supervised-learning oracle inequality with self-mitigating statistical error.
-
On Least Squares Estimation under Heteroscedastic and Heavy-Tailed Errors
Under finite moments and a local envelope growth condition, the least squares estimator in nonparametric regression can achieve minimax rates with heavy-tailed, covariate-dependent errors.
-
Minimum Norm Interpolation via The Local Theory of Banach Spaces: The Role of Gaussianity
The sharp MSE bound for the ℓ1-minimum-norm interpolator under isotropic Gaussian covariates is recovered via the geometry of symmetric Gaussian polytopes, without the convex Gaussian min-max theorem.
-
Double Preconditioning (DoPr): Optimization for Test-Time Performance, not Validation Loss
Double preconditioning (DoPr) improves downstream task performance in test-time feedback settings without consistent gains in validation loss.
-
Statistical guarantees for continuous-time policy evaluation: blessing of ellipticity and new tradeoffs
For continuous-time policy evaluation, the LSTD estimator's H1 error scales as the square root of (approximation error plus m/T), with a trajectory length that can be nearly linear in the number of basis functions whe...
-
On the Efficiency of ERM in Feature Learning
ERM over a union of linear feature classes achieves excess risk within a factor of two of the oracle that knows the optimal feature map, asymptotically, under size conditions on the feature set.
Discussion (0). Continue with ORCID to comment.