REVIEW 12 cited by
Dataset Condensation with Gradient Matching
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
As the state-of-the-art machine learning methods in many fields rely on larger datasets, storing datasets and training models on them become significantly more expensive. This paper proposes a training set synthesis technique for data-efficient learning, called Dataset Condensation, that learns to condense large dataset into a small set of informative synthetic samples for training deep neural networks from scratch. We formulate this goal as a gradient matching problem between the gradients of deep neural network weights that are trained on the original and our synthetic data. We rigorously evaluate its performance in several computer vision benchmarks and demonstrate that it significantly outperforms the state-of-the-art methods. Finally we explore the use of our method in continual learning and neural architecture search and report promising gains when limited memory and computations are available.
Forward citations
Cited by 12 Pith papers
-
SoK: Can Synthetic Images Replace Real Data? A Survey of Utility and Privacy of Synthetic Image Generation
A systematic survey and benchmark showing that diffusion-based synthetic data can achieve better utility-privacy tradeoffs than DP-SGD on real data for some image classifiers, with the best release strategy depending ...
-
Hard Labels In! Rethinking the Role of Hard Labels in Mitigating Local Semantic Drift
A soft-hard-soft training schedule uses hard labels as an intermediate anchor to correct local semantic drift and improves accuracy under 100x-reduced soft-label storage.
-
A computational fluid dynamics model for the simulation of flashboiling flow inside pressurized metered dose inhalers
The abstract claims a first-of-kind open-source CFD model, combining Volume-of-Fluid and cavitation modeling, that quantitatively predicts flashboiling flow in pressurized metered dose inhalers, but the supplied text ...
-
GVD: Guiding Video Diffusion Model for Scalable Video Distillation
GVD guides a pre-trained video diffusion model with clustering-derived features to distill video datasets, outperforming prior methods on MiniUCF and HMDB51 while retaining over 70% of full-data accuracy using under 4...
-
Approximating Language Model Training Data from Weights
A gradient-based greedy selection method (SELECT) recovers effective substitute fine-tuning data from two language model checkpoints, approaching the original model's performance on classification and SFT tasks.
-
Reasoning Like an Economist: Post-Training on Economic Problems Induces Strategic Generalization in LLMs
Post-training an LLM on economic reasoning problems improves accuracy on economic benchmarks and, without game-specific training, raises its Nash equilibrium frequency and win rates in strategic games.
-
GLOBE: Trajectory-Aligned Gradient Matching with Structured SparseOptimization for Coreset Selection
GLOBE selects compact training subsets by matching multi-checkpoint gradient trajectories and their second-order statistics under structured sparsity, outperforming prior coreset methods on six image benchmarks.
-
NMS: Efficient Edge DNN Training via Near-Memory Sampling on Manifolds
A t-SNE-based sampling algorithm with differential evolution and near-memory hardware is claimed to speed up edge DNN training and reduce memory energy.
-
Data-Distill-Net: A Data Distillation Approach Tailored for Reply-based Continual Learning
A plug-in module that generates learned soft labels for memory buffer samples improves accuracy and reduces forgetting across several replay-based continual learning baselines.
-
Data-Efficient Ensemble Weather Forecasting with Diffusion Models
Training an autoregressive diffusion weather forecaster on 20% of ERA5 data selected uniformly by calendar month matches full-data CRPS/RMSE and improves the spread-skill ratio on the 2018 test year.
-
Domain-Generalization to Improve Learning in Meta-Learning Algorithms
DGS-MAML layers gradient matching onto SharpMAML and claims O(1/T) convergence and tighter PAC-Bayes bounds, but the displayed theorems give O(1/sqrt T) under the paper's own parameter choices.
-
Leveraging Distribution Matching to Make Approximate Machine Unlearning Faster
A dual data and loss-centric method claims to speed up machine unlearning, but its MIA regularizer cancels itself and the test set is leaked into training.
Discussion (0). Sign in to comment.