Soft Label Pruning and Quantization for Large-Scale Dataset Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | Lingao, Xiao, He, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Are Large-scale Soft Labels Necessary for Large-scale Dataset Distillation?
by: Xiao, Lingao, et al.
Published: (2024)
by: Xiao, Lingao, et al.
Published: (2024)
Rethinking Large-scale Dataset Compression: Shifting Focus From Labels to Images
by: Xiao, Lingao, et al.
Published: (2025)
by: Xiao, Lingao, et al.
Published: (2025)
Training-Free Dataset Pruning for Instance Segmentation
by: Dai, Yalun, et al.
Published: (2025)
by: Dai, Yalun, et al.
Published: (2025)
Dataset Color Quantization: A Training-Oriented Framework for Dataset-Level Compression
by: Yu, Chenyue, et al.
Published: (2026)
by: Yu, Chenyue, et al.
Published: (2026)
Distill the Best, Ignore the Rest: Improving Dataset Distillation with Loss-Value-Based Pruning
by: Moser, Brian B., et al.
Published: (2024)
by: Moser, Brian B., et al.
Published: (2024)
Accelerating Large-Scale Dataset Distillation via Exploration-Exploitation Optimization
by: Alahmadi, Muhammad J., et al.
Published: (2026)
by: Alahmadi, Muhammad J., et al.
Published: (2026)
Distill Gold from Massive Ores: Bi-level Data Pruning towards Efficient Dataset Distillation
by: Xu, Yue, et al.
Published: (2023)
by: Xu, Yue, et al.
Published: (2023)
Multi-Label Knowledge Distillation
by: Yang, Penghui, et al.
Published: (2023)
by: Yang, Penghui, et al.
Published: (2023)
On the Diversity and Realism of Distilled Dataset: An Efficient Dataset Distillation Paradigm
by: Sun, Peng, et al.
Published: (2023)
by: Sun, Peng, et al.
Published: (2023)
Dataset Distillation via the Wasserstein Metric
by: Liu, Haoyang, et al.
Published: (2023)
by: Liu, Haoyang, et al.
Published: (2023)
Dataset Distillation with Neural Characteristic Function: A Minmax Perspective
by: Wang, Shaobo, et al.
Published: (2025)
by: Wang, Shaobo, et al.
Published: (2025)
Self-Supervised Quantization-Aware Knowledge Distillation
by: Zhao, Kaiqi, et al.
Published: (2024)
by: Zhao, Kaiqi, et al.
Published: (2024)
EPSD: Early Pruning with Self-Distillation for Efficient Model Compression
by: Chen, Dong, et al.
Published: (2024)
by: Chen, Dong, et al.
Published: (2024)
Scale Efficient Training for Large Datasets
by: Zhou, Qing, et al.
Published: (2025)
by: Zhou, Qing, et al.
Published: (2025)
Hyperbolic Dataset Distillation
by: Li, Wenyuan, et al.
Published: (2025)
by: Li, Wenyuan, et al.
Published: (2025)
Dataset Distillation via Curriculum Data Synthesis in Large Data Era
by: Yin, Zeyuan, et al.
Published: (2023)
by: Yin, Zeyuan, et al.
Published: (2023)
Small Scale Data-Free Knowledge Distillation
by: Liu, He, et al.
Published: (2024)
by: Liu, He, et al.
Published: (2024)
Automatic Joint Structured Pruning and Quantization for Efficient Neural Network Training and Compression
by: Qu, Xiaoyi, et al.
Published: (2025)
by: Qu, Xiaoyi, et al.
Published: (2025)
Generative Dataset Distillation Based on Self-knowledge Distillation
by: Li, Longzhen, et al.
Published: (2025)
by: Li, Longzhen, et al.
Published: (2025)
Mitigating Bias in Dataset Distillation
by: Cui, Justin, et al.
Published: (2024)
by: Cui, Justin, et al.
Published: (2024)
Distilling Dataset into Neural Field
by: Shin, Donghyeok, et al.
Published: (2025)
by: Shin, Donghyeok, et al.
Published: (2025)
Aneumo: A Large-Scale Comprehensive Synthetic Dataset of Aneurysm Hemodynamics
by: Li, Xigui, et al.
Published: (2025)
by: Li, Xigui, et al.
Published: (2025)
FairDD: Fair Dataset Distillation
by: Zhou, Qihang, et al.
Published: (2024)
by: Zhou, Qihang, et al.
Published: (2024)
DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models
by: Alvar, Saeed Ranjbar, et al.
Published: (2025)
by: Alvar, Saeed Ranjbar, et al.
Published: (2025)
Jaccard Metric Losses: Optimizing the Jaccard Index with Soft Labels
by: Wang, Zifu, et al.
Published: (2023)
by: Wang, Zifu, et al.
Published: (2023)
Dice Semimetric Losses: Optimizing the Dice Score with Soft Labels
by: Wang, Zifu, et al.
Published: (2023)
by: Wang, Zifu, et al.
Published: (2023)
Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation
by: Xie, Jingjing, et al.
Published: (2024)
by: Xie, Jingjing, et al.
Published: (2024)
Unlocking Dataset Distillation with Diffusion Models
by: Moser, Brian B., et al.
Published: (2024)
by: Moser, Brian B., et al.
Published: (2024)
Importance-Aware Adaptive Dataset Distillation
by: Li, Guang, et al.
Published: (2024)
by: Li, Guang, et al.
Published: (2024)
PHI-S: Distribution Balancing for Label-Free Multi-Teacher Distillation
by: Ranzinger, Mike, et al.
Published: (2024)
by: Ranzinger, Mike, et al.
Published: (2024)
Kaputt: A Large-Scale Dataset for Visual Defect Detection
by: Höfer, Sebastian, et al.
Published: (2025)
by: Höfer, Sebastian, et al.
Published: (2025)
Taming Diffusion for Dataset Distillation with High Representativeness
by: Zhao, Lin, et al.
Published: (2025)
by: Zhao, Lin, et al.
Published: (2025)
Generative Dataset Distillation Based on Diffusion Model
by: Su, Duo, et al.
Published: (2024)
by: Su, Duo, et al.
Published: (2024)
SDXL-Lightning: Progressive Adversarial Diffusion Distillation
by: Lin, Shanchuan, et al.
Published: (2024)
by: Lin, Shanchuan, et al.
Published: (2024)
SAS: Semantic-aware Sampling for Generative Dataset Distillation
by: Li, Mingzhuo, et al.
Published: (2026)
by: Li, Mingzhuo, et al.
Published: (2026)
A Study in Dataset Distillation for Image Super-Resolution
by: Dietz, Tobias, et al.
Published: (2025)
by: Dietz, Tobias, et al.
Published: (2025)
PRISM: Diversifying Dataset Distillation by Decoupling Architectural Priors
by: Moser, Brian B., et al.
Published: (2025)
by: Moser, Brian B., et al.
Published: (2025)
Group Distributionally Robust Dataset Distillation with Risk Minimization
by: Vahidian, Saeed, et al.
Published: (2024)
by: Vahidian, Saeed, et al.
Published: (2024)
Rethinking Dataset Distillation: Hard Truths about Soft Labels
by: Dey, Priyam, et al.
Published: (2026)
by: Dey, Priyam, et al.
Published: (2026)
SearchAD: Large-Scale Rare Image Retrieval Dataset for Autonomous Driving
by: Embacher, Felix, et al.
Published: (2026)
by: Embacher, Felix, et al.
Published: (2026)
Similar Items
-
Are Large-scale Soft Labels Necessary for Large-scale Dataset Distillation?
by: Xiao, Lingao, et al.
Published: (2024) -
Rethinking Large-scale Dataset Compression: Shifting Focus From Labels to Images
by: Xiao, Lingao, et al.
Published: (2025) -
Training-Free Dataset Pruning for Instance Segmentation
by: Dai, Yalun, et al.
Published: (2025) -
Dataset Color Quantization: A Training-Oriented Framework for Dataset-Level Compression
by: Yu, Chenyue, et al.
Published: (2026) -
Distill the Best, Ignore the Rest: Improving Dataset Distillation with Loss-Value-Based Pruning
by: Moser, Brian B., et al.
Published: (2024)