Towards Principled Dataset Distillation: A Spectral Distribution Perspective
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Ruixi, Wang, Shaobo, Chen, Jiahuan, Liu, Zhiyuan, Yang, Yicun, Chen, Zhaorun, Li, Zekai, Li, Kaixin, Wang, Xinming, Yi, Hongzhu, Wang, Kai, Zhang, Linfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Grounding and Enhancing Informativeness and Utility in Dataset Distillation
von: Wang, Shaobo, et al.
Veröffentlicht: (2026)
von: Wang, Shaobo, et al.
Veröffentlicht: (2026)
Dataset Distillation with Neural Characteristic Function: A Minmax Perspective
von: Wang, Shaobo, et al.
Veröffentlicht: (2025)
von: Wang, Shaobo, et al.
Veröffentlicht: (2025)
Not All Samples Should Be Utilized Equally: Towards Understanding and Improving Dataset Distillation
von: Wang, Shaobo, et al.
Veröffentlicht: (2024)
von: Wang, Shaobo, et al.
Veröffentlicht: (2024)
dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2025)
Efficient Multi-modal Large Language Models via Progressive Consistency Distillation
von: Wen, Zichen, et al.
Veröffentlicht: (2025)
von: Wen, Zichen, et al.
Veröffentlicht: (2025)
Diffusion LLM with Native Variable Generation Lengths: Let [EOS] Lead the Way
von: Yang, Yicun, et al.
Veröffentlicht: (2025)
von: Yang, Yicun, et al.
Veröffentlicht: (2025)
Dynamic Deep Graph Learning for Incomplete Multi-View Clustering with Masked Graph Reconstruction Loss
von: Zhang, Zhenghao, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenghao, et al.
Veröffentlicht: (2025)
The Devil behind the mask: An emergent safety vulnerability of Diffusion LLMs
von: Wen, Zichen, et al.
Veröffentlicht: (2025)
von: Wen, Zichen, et al.
Veröffentlicht: (2025)
UNSEEN: Enhancing Dataset Pruning from a Generalization Perspective
von: Xu, Furui, et al.
Veröffentlicht: (2025)
von: Xu, Furui, et al.
Veröffentlicht: (2025)
RPO:Reinforcement Fine-Tuning with Partial Reasoning Optimization
von: Yi, Hongzhu, et al.
Veröffentlicht: (2026)
von: Yi, Hongzhu, et al.
Veröffentlicht: (2026)
Multimodal Video Emotion Recognition with Reliable Reasoning Priors
von: Wang, Zhepeng, et al.
Veröffentlicht: (2025)
von: Wang, Zhepeng, et al.
Veröffentlicht: (2025)
Flash-Unified: A Training-Free and Task-Aware Acceleration Framework for Native Unified Models
von: Ke, Junlong, et al.
Veröffentlicht: (2026)
von: Ke, Junlong, et al.
Veröffentlicht: (2026)
Prioritize Alignment in Dataset Distillation
von: Li, Zekai, et al.
Veröffentlicht: (2024)
von: Li, Zekai, et al.
Veröffentlicht: (2024)
SpeCa: Accelerating Diffusion Transformers with Speculative Feature Caching
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)
Emphasizing Discriminative Features for Dataset Distillation in Complex Scenarios
von: Wang, Kai, et al.
Veröffentlicht: (2024)
von: Wang, Kai, et al.
Veröffentlicht: (2024)
Four Eyes Are Better Than Two: Harnessing the Collaborative Potential of Large Models via Differentiated Thinking and Complementary Ensembles
von: Xie, Jun, et al.
Veröffentlicht: (2025)
von: Xie, Jun, et al.
Veröffentlicht: (2025)
Team of One: Cracking Complex Video QA with Model Synergy
von: Xie, Jun, et al.
Veröffentlicht: (2025)
von: Xie, Jun, et al.
Veröffentlicht: (2025)
Towards Lossless Dataset Distillation via Difficulty-Aligned Trajectory Matching
von: Guo, Ziyao, et al.
Veröffentlicht: (2023)
von: Guo, Ziyao, et al.
Veröffentlicht: (2023)
DD-Ranking: Rethinking the Evaluation of Dataset Distillation
von: Li, Zekai, et al.
Veröffentlicht: (2025)
von: Li, Zekai, et al.
Veröffentlicht: (2025)
Transcriptome analysis reveals important regulatory factors for condensed tannins synthesis in acorn
von: Liwen Wu, et al.
Veröffentlicht: (2024)
von: Liwen Wu, et al.
Veröffentlicht: (2024)
DDTime: Dataset Distillation with Spectral Alignment and Information Bottleneck for Time-Series Forecasting
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
Understanding Dataset Distillation via Spectral Filtering
von: Bo, Deyu, et al.
Veröffentlicht: (2025)
von: Bo, Deyu, et al.
Veröffentlicht: (2025)
Cer-Eval: Certifiable and Cost-Efficient Evaluation Framework for LLMs
von: Wang, Ganghua, et al.
Veröffentlicht: (2025)
von: Wang, Ganghua, et al.
Veröffentlicht: (2025)
OPUS: Towards Efficient and Principled Data Selection in Large Language Model Pre-training in Every Iteration
von: Wang, Shaobo, et al.
Veröffentlicht: (2026)
von: Wang, Shaobo, et al.
Veröffentlicht: (2026)
Towards Adversarially Robust Dataset Distillation by Curvature Regularization
von: Xue, Eric, et al.
Veröffentlicht: (2024)
von: Xue, Eric, et al.
Veröffentlicht: (2024)
Lower Bound of Nodal Sets in Elliptic Homogenization and Functions with Strong Maximum Principle
von: Li, Jiahuan, et al.
Veröffentlicht: (2025)
von: Li, Jiahuan, et al.
Veröffentlicht: (2025)
IMS3: Breaking Distributional Aggregation in Diffusion-Based Dataset Distillation
von: Wang, Chenru, et al.
Veröffentlicht: (2026)
von: Wang, Chenru, et al.
Veröffentlicht: (2026)
Data Whisperer: Efficient Data Selection for Task-Specific LLM Fine-Tuning via Few-Shot In-Context Learning
von: Wang, Shaobo, et al.
Veröffentlicht: (2025)
von: Wang, Shaobo, et al.
Veröffentlicht: (2025)
Group Distributionally Robust Dataset Distillation with Risk Minimization
von: Vahidian, Saeed, et al.
Veröffentlicht: (2024)
von: Vahidian, Saeed, et al.
Veröffentlicht: (2024)
Distilled Large Language Model-Driven Dynamic Sparse Expert Activation Mechanism
von: Chen, Qinghui, et al.
Veröffentlicht: (2026)
von: Chen, Qinghui, et al.
Veröffentlicht: (2026)
Toward Medical Deepfake Detection: A Comprehensive Dataset and Novel Method
von: Li, Shuaibo, et al.
Veröffentlicht: (2025)
von: Li, Shuaibo, et al.
Veröffentlicht: (2025)
Distilling Long-tailed Datasets
von: Zhao, Zhenghao, et al.
Veröffentlicht: (2024)
von: Zhao, Zhenghao, et al.
Veröffentlicht: (2024)
ShieldAgent: Shielding Agents via Verifiable Safety Policy Reasoning
von: Chen, Zhaorun, et al.
Veröffentlicht: (2025)
von: Chen, Zhaorun, et al.
Veröffentlicht: (2025)
1.x-Distill: Breaking the Diversity, Quality, and Efficiency Barrier in Distribution Matching Distillation
von: Li, Haoyu, et al.
Veröffentlicht: (2026)
von: Li, Haoyu, et al.
Veröffentlicht: (2026)
MT-PATCHER: Selective and Extendable Knowledge Distillation from Large Language Models for Machine Translation
von: Li, Jiahuan, et al.
Veröffentlicht: (2024)
von: Li, Jiahuan, et al.
Veröffentlicht: (2024)
Distribution-aware Dataset Distillation for Efficient Image Restoration
von: Zheng, Zhuoran, et al.
Veröffentlicht: (2025)
von: Zheng, Zhuoran, et al.
Veröffentlicht: (2025)
dVLA: Diffusion Vision-Language-Action Model with Multimodal Chain-of-Thought
von: Wen, Junjie, et al.
Veröffentlicht: (2025)
von: Wen, Junjie, et al.
Veröffentlicht: (2025)
FreeFly-Thinking : Aligning Chain-of-Thought Reasoning with Continuous UAV Navigation
von: Zhou, Jiaxu, et al.
Veröffentlicht: (2026)
von: Zhou, Jiaxu, et al.
Veröffentlicht: (2026)
Overexpression of Anthocyanidin Reductase Increases Flavonoids Content to Combat Fusarium Wilt in the Root Xylem of Vernicia montana
von: Jia Wang, et al.
Veröffentlicht: (2026)
von: Jia Wang, et al.
Veröffentlicht: (2026)
Distill Video Datasets into Images
von: Zhao, Zhenghao, et al.
Veröffentlicht: (2025)
von: Zhao, Zhenghao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Grounding and Enhancing Informativeness and Utility in Dataset Distillation
von: Wang, Shaobo, et al.
Veröffentlicht: (2026) -
Dataset Distillation with Neural Characteristic Function: A Minmax Perspective
von: Wang, Shaobo, et al.
Veröffentlicht: (2025) -
Not All Samples Should Be Utilized Equally: Towards Understanding and Improving Dataset Distillation
von: Wang, Shaobo, et al.
Veröffentlicht: (2024) -
dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2025) -
Efficient Multi-modal Large Language Models via Progressive Consistency Distillation
von: Wen, Zichen, et al.
Veröffentlicht: (2025)