Towards Principled Dataset Distillation: A Spectral Distribution Perspective
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Ruixi, Wang, Shaobo, Chen, Jiahuan, Liu, Zhiyuan, Yang, Yicun, Chen, Zhaorun, Li, Zekai, Li, Kaixin, Wang, Xinming, Yi, Hongzhu, Wang, Kai, Zhang, Linfeng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Grounding and Enhancing Informativeness and Utility in Dataset Distillation
by: Wang, Shaobo, et al.
Published: (2026)
by: Wang, Shaobo, et al.
Published: (2026)
Dataset Distillation with Neural Characteristic Function: A Minmax Perspective
by: Wang, Shaobo, et al.
Published: (2025)
by: Wang, Shaobo, et al.
Published: (2025)
Not All Samples Should Be Utilized Equally: Towards Understanding and Improving Dataset Distillation
by: Wang, Shaobo, et al.
Published: (2024)
by: Wang, Shaobo, et al.
Published: (2024)
dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching
by: Liu, Zhiyuan, et al.
Published: (2025)
by: Liu, Zhiyuan, et al.
Published: (2025)
Efficient Multi-modal Large Language Models via Progressive Consistency Distillation
by: Wen, Zichen, et al.
Published: (2025)
by: Wen, Zichen, et al.
Published: (2025)
Diffusion LLM with Native Variable Generation Lengths: Let [EOS] Lead the Way
by: Yang, Yicun, et al.
Published: (2025)
by: Yang, Yicun, et al.
Published: (2025)
Dynamic Deep Graph Learning for Incomplete Multi-View Clustering with Masked Graph Reconstruction Loss
by: Zhang, Zhenghao, et al.
Published: (2025)
by: Zhang, Zhenghao, et al.
Published: (2025)
The Devil behind the mask: An emergent safety vulnerability of Diffusion LLMs
by: Wen, Zichen, et al.
Published: (2025)
by: Wen, Zichen, et al.
Published: (2025)
UNSEEN: Enhancing Dataset Pruning from a Generalization Perspective
by: Xu, Furui, et al.
Published: (2025)
by: Xu, Furui, et al.
Published: (2025)
RPO:Reinforcement Fine-Tuning with Partial Reasoning Optimization
by: Yi, Hongzhu, et al.
Published: (2026)
by: Yi, Hongzhu, et al.
Published: (2026)
Multimodal Video Emotion Recognition with Reliable Reasoning Priors
by: Wang, Zhepeng, et al.
Published: (2025)
by: Wang, Zhepeng, et al.
Published: (2025)
Flash-Unified: A Training-Free and Task-Aware Acceleration Framework for Native Unified Models
by: Ke, Junlong, et al.
Published: (2026)
by: Ke, Junlong, et al.
Published: (2026)
Prioritize Alignment in Dataset Distillation
by: Li, Zekai, et al.
Published: (2024)
by: Li, Zekai, et al.
Published: (2024)
SpeCa: Accelerating Diffusion Transformers with Speculative Feature Caching
by: Liu, Jiacheng, et al.
Published: (2025)
by: Liu, Jiacheng, et al.
Published: (2025)
Emphasizing Discriminative Features for Dataset Distillation in Complex Scenarios
by: Wang, Kai, et al.
Published: (2024)
by: Wang, Kai, et al.
Published: (2024)
Four Eyes Are Better Than Two: Harnessing the Collaborative Potential of Large Models via Differentiated Thinking and Complementary Ensembles
by: Xie, Jun, et al.
Published: (2025)
by: Xie, Jun, et al.
Published: (2025)
Team of One: Cracking Complex Video QA with Model Synergy
by: Xie, Jun, et al.
Published: (2025)
by: Xie, Jun, et al.
Published: (2025)
Towards Lossless Dataset Distillation via Difficulty-Aligned Trajectory Matching
by: Guo, Ziyao, et al.
Published: (2023)
by: Guo, Ziyao, et al.
Published: (2023)
DD-Ranking: Rethinking the Evaluation of Dataset Distillation
by: Li, Zekai, et al.
Published: (2025)
by: Li, Zekai, et al.
Published: (2025)
Transcriptome analysis reveals important regulatory factors for condensed tannins synthesis in acorn
by: Liwen Wu, et al.
Published: (2024)
by: Liwen Wu, et al.
Published: (2024)
DDTime: Dataset Distillation with Spectral Alignment and Information Bottleneck for Time-Series Forecasting
by: Li, Yuqi, et al.
Published: (2025)
by: Li, Yuqi, et al.
Published: (2025)
Understanding Dataset Distillation via Spectral Filtering
by: Bo, Deyu, et al.
Published: (2025)
by: Bo, Deyu, et al.
Published: (2025)
Cer-Eval: Certifiable and Cost-Efficient Evaluation Framework for LLMs
by: Wang, Ganghua, et al.
Published: (2025)
by: Wang, Ganghua, et al.
Published: (2025)
OPUS: Towards Efficient and Principled Data Selection in Large Language Model Pre-training in Every Iteration
by: Wang, Shaobo, et al.
Published: (2026)
by: Wang, Shaobo, et al.
Published: (2026)
Towards Adversarially Robust Dataset Distillation by Curvature Regularization
by: Xue, Eric, et al.
Published: (2024)
by: Xue, Eric, et al.
Published: (2024)
Lower Bound of Nodal Sets in Elliptic Homogenization and Functions with Strong Maximum Principle
by: Li, Jiahuan, et al.
Published: (2025)
by: Li, Jiahuan, et al.
Published: (2025)
IMS3: Breaking Distributional Aggregation in Diffusion-Based Dataset Distillation
by: Wang, Chenru, et al.
Published: (2026)
by: Wang, Chenru, et al.
Published: (2026)
Data Whisperer: Efficient Data Selection for Task-Specific LLM Fine-Tuning via Few-Shot In-Context Learning
by: Wang, Shaobo, et al.
Published: (2025)
by: Wang, Shaobo, et al.
Published: (2025)
Group Distributionally Robust Dataset Distillation with Risk Minimization
by: Vahidian, Saeed, et al.
Published: (2024)
by: Vahidian, Saeed, et al.
Published: (2024)
Distilled Large Language Model-Driven Dynamic Sparse Expert Activation Mechanism
by: Chen, Qinghui, et al.
Published: (2026)
by: Chen, Qinghui, et al.
Published: (2026)
Toward Medical Deepfake Detection: A Comprehensive Dataset and Novel Method
by: Li, Shuaibo, et al.
Published: (2025)
by: Li, Shuaibo, et al.
Published: (2025)
Distilling Long-tailed Datasets
by: Zhao, Zhenghao, et al.
Published: (2024)
by: Zhao, Zhenghao, et al.
Published: (2024)
ShieldAgent: Shielding Agents via Verifiable Safety Policy Reasoning
by: Chen, Zhaorun, et al.
Published: (2025)
by: Chen, Zhaorun, et al.
Published: (2025)
1.x-Distill: Breaking the Diversity, Quality, and Efficiency Barrier in Distribution Matching Distillation
by: Li, Haoyu, et al.
Published: (2026)
by: Li, Haoyu, et al.
Published: (2026)
MT-PATCHER: Selective and Extendable Knowledge Distillation from Large Language Models for Machine Translation
by: Li, Jiahuan, et al.
Published: (2024)
by: Li, Jiahuan, et al.
Published: (2024)
Distribution-aware Dataset Distillation for Efficient Image Restoration
by: Zheng, Zhuoran, et al.
Published: (2025)
by: Zheng, Zhuoran, et al.
Published: (2025)
dVLA: Diffusion Vision-Language-Action Model with Multimodal Chain-of-Thought
by: Wen, Junjie, et al.
Published: (2025)
by: Wen, Junjie, et al.
Published: (2025)
FreeFly-Thinking : Aligning Chain-of-Thought Reasoning with Continuous UAV Navigation
by: Zhou, Jiaxu, et al.
Published: (2026)
by: Zhou, Jiaxu, et al.
Published: (2026)
Overexpression of Anthocyanidin Reductase Increases Flavonoids Content to Combat Fusarium Wilt in the Root Xylem of Vernicia montana
by: Jia Wang, et al.
Published: (2026)
by: Jia Wang, et al.
Published: (2026)
Distill Video Datasets into Images
by: Zhao, Zhenghao, et al.
Published: (2025)
by: Zhao, Zhenghao, et al.
Published: (2025)
Similar Items
-
Grounding and Enhancing Informativeness and Utility in Dataset Distillation
by: Wang, Shaobo, et al.
Published: (2026) -
Dataset Distillation with Neural Characteristic Function: A Minmax Perspective
by: Wang, Shaobo, et al.
Published: (2025) -
Not All Samples Should Be Utilized Equally: Towards Understanding and Improving Dataset Distillation
by: Wang, Shaobo, et al.
Published: (2024) -
dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching
by: Liu, Zhiyuan, et al.
Published: (2025) -
Efficient Multi-modal Large Language Models via Progressive Consistency Distillation
by: Wen, Zichen, et al.
Published: (2025)