SOTAVerified

Dataset Distillation

Dataset distillation is the task of synthesizing a small dataset such that models trained on it achieve high performance on the original large dataset. A dataset distillation algorithm takes as input a large real dataset to be distilled (training set), and outputs a small synthetic distilled dataset, which is evaluated via testing models trained on this distilled dataset on a separate real dataset (validation/test set). A good small distilled dataset is not only useful in dataset understanding, but has various applications (e.g., continual learning, privacy, neural architecture search, etc.).

Papers

Showing 125 of 216 papers

TitleStatusHype
Information-Guided Diffusion Sampling for Dataset Distillation0
Task-Specific Generative Dataset Distillation with Difficulty-Guided SamplingCode0
Dataset Distillation via Vision-Language Category PrototypeCode1
FADRM: Fast and Accurate Data Residual Matching for Dataset DistillationCode1
CaO_2: Rectifying Inconsistencies in Diffusion-Based Dataset DistillationCode1
FedWSIDD: Federated Whole Slide Image Classification via Dataset Distillation0
Dataset distillation for memorized data: Soft labels can leak held-out teacher knowledgeCode0
Flowing Datasets with Wasserstein over Wasserstein Gradient FlowsCode1
OD3: Optimization-free Dataset Distillation for Object DetectionCode1
Hyperbolic Dataset DistillationCode0
Data-Distill-Net: A Data Distillation Approach Tailored for Reply-based Continual Learning0
Diversity-Driven Generative Dataset Distillation Based on Diffusion Model with Self-Adaptive Memory0
MGD^3: Mode-Guided Dataset Distillation using Diffusion Models0
Taming Diffusion for Dataset Distillation with High RepresentativenessCode1
CONCORD: Concept-Informed Diffusion for Dataset DistillationCode0
Contrastive Learning-Enhanced Trajectory Matching for Small-Scale Dataset Distillation0
Exploring Generalized Gait Recognition: Reducing Redundancy and Noise within Indoor and Outdoor DatasetsCode0
DD-Ranking: Rethinking the Evaluation of Dataset DistillationCode2
Beyond Modality Collapse: Representations Blending for Multimodal Dataset Distillation0
Leveraging Multi-Modal Information to Enhance Dataset Distillation0
Dataset Distillation with Probabilistic Latent Features0
Video Dataset Condensation with Diffusion Models0
UniDetox: Universal Detoxification of Large Language Models via Dataset DistillationCode0
Latent Video Dataset Distillation0
Distribution-aware Dataset Distillation for Efficient Image Restoration0
Show:102550
← PrevPage 1 of 9Next →

No leaderboard results yet.