← back to paper
arxiv: 2506.20025 · 2 revisions
Thumb on the Scale: Optimal Loss Weighting in Last Layer Retraining