Grounding and Enhancing Informativeness and Utility in Dataset Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Shaobo, Yang, Yantai, Chen, Guo, Li, Peiru, Li, Kaixin, Zhou, Yufa, Chen, Zhaorun, Zhang, Linfeng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Not All Samples Should Be Utilized Equally: Towards Understanding and Improving Dataset Distillation
by: Wang, Shaobo, et al.
Published: (2024)
by: Wang, Shaobo, et al.
Published: (2024)
DRUPI: Dataset Reduction Using Privileged Information
by: Wang, Shaobo, et al.
Published: (2024)
by: Wang, Shaobo, et al.
Published: (2024)
Dataset Distillation with Neural Characteristic Function: A Minmax Perspective
by: Wang, Shaobo, et al.
Published: (2025)
by: Wang, Shaobo, et al.
Published: (2025)
Towards Principled Dataset Distillation: A Spectral Distribution Perspective
by: Wu, Ruixi, et al.
Published: (2026)
by: Wu, Ruixi, et al.
Published: (2026)
SpeCa: Accelerating Diffusion Transformers with Speculative Feature Caching
by: Liu, Jiacheng, et al.
Published: (2025)
by: Liu, Jiacheng, et al.
Published: (2025)
Prioritize Alignment in Dataset Distillation
by: Li, Zekai, et al.
Published: (2024)
by: Li, Zekai, et al.
Published: (2024)
Why Do Transformers Fail to Forecast Time Series In-Context?
by: Zhou, Yufa, et al.
Published: (2025)
by: Zhou, Yufa, et al.
Published: (2025)
Cer-Eval: Certifiable and Cost-Efficient Evaluation Framework for LLMs
by: Wang, Ganghua, et al.
Published: (2025)
by: Wang, Ganghua, et al.
Published: (2025)
Agentic Proposing: Enhancing Large Language Model Reasoning via Compositional Skill Synthesis
by: Jiao, Zhengbo, et al.
Published: (2026)
by: Jiao, Zhengbo, et al.
Published: (2026)
DDTime: Dataset Distillation with Spectral Alignment and Information Bottleneck for Time-Series Forecasting
by: Li, Yuqi, et al.
Published: (2025)
by: Li, Yuqi, et al.
Published: (2025)
Information-Guided Diffusion Sampling for Dataset Distillation
by: Ye, Linfeng, et al.
Published: (2025)
by: Ye, Linfeng, et al.
Published: (2025)
Enhancing Vision-Language Model Reliability with Uncertainty-Guided Dropout Decoding
by: Fang, Yixiong, et al.
Published: (2024)
by: Fang, Yixiong, et al.
Published: (2024)
ARMs: Adaptive Red-Teaming Agent against Multimodal Models with Plug-and-Play Attacks
by: Chen, Zhaorun, et al.
Published: (2025)
by: Chen, Zhaorun, et al.
Published: (2025)
Task-Specific Generative Dataset Distillation with Difficulty-Guided Sampling
by: Li, Mingzhuo, et al.
Published: (2025)
by: Li, Mingzhuo, et al.
Published: (2025)
Towards Mitigating Architecture Overfitting on Distilled Datasets
by: Zhong, Xuyang, et al.
Published: (2023)
by: Zhong, Xuyang, et al.
Published: (2023)
SAS: Semantic-aware Sampling for Generative Dataset Distillation
by: Li, Mingzhuo, et al.
Published: (2026)
by: Li, Mingzhuo, et al.
Published: (2026)
Distilling a Small Utility-Based Passage Selector to Enhance Retrieval-Augmented Generation
by: Zhang, Hengran, et al.
Published: (2025)
by: Zhang, Hengran, et al.
Published: (2025)
Turning Black Box into White Box: Dataset Distillation Leaks
by: Chen, Huajie, et al.
Published: (2026)
by: Chen, Huajie, et al.
Published: (2026)
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
by: Chen, Xiaoyu, et al.
Published: (2025)
by: Chen, Xiaoyu, et al.
Published: (2025)
dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching
by: Liu, Zhiyuan, et al.
Published: (2025)
by: Liu, Zhiyuan, et al.
Published: (2025)
Group Relative Knowledge Distillation: Learning from Teacher's Relational Inductive Bias
by: Li, Chao, et al.
Published: (2025)
by: Li, Chao, et al.
Published: (2025)
Path-Guided Flow Matching for Dataset Distillation
by: Li, Xuhui, et al.
Published: (2026)
by: Li, Xuhui, et al.
Published: (2026)
scDD: Latent Codes Based scRNA-seq Dataset Distillation with Foundation Model Knowledge
by: Yu, Zhen, et al.
Published: (2025)
by: Yu, Zhen, et al.
Published: (2025)
Enhancing Graph Neural Networks with Limited Labeled Data by Actively Distilling Knowledge from Large Language Models
by: Li, Quan, et al.
Published: (2024)
by: Li, Quan, et al.
Published: (2024)
Online Adversarial Knowledge Distillation for Graph Neural Networks
by: Wang, Can, et al.
Published: (2021)
by: Wang, Can, et al.
Published: (2021)
Enhancing Knowledge Graph Completion with GNN Distillation and Probabilistic Interaction Modeling
by: Wang, Lingzhi, et al.
Published: (2025)
by: Wang, Lingzhi, et al.
Published: (2025)
Dataset Distillation-based Hybrid Federated Learning on Non-IID Data
by: Shi, Xiufang, et al.
Published: (2024)
by: Shi, Xiufang, et al.
Published: (2024)
Solving Continual Offline Reinforcement Learning with Decision Transformer
by: Huang, Kaixin, et al.
Published: (2024)
by: Huang, Kaixin, et al.
Published: (2024)
QuickDrop: Efficient Federated Unlearning by Integrated Dataset Distillation
by: Dhasade, Akash, et al.
Published: (2023)
by: Dhasade, Akash, et al.
Published: (2023)
Hyperbolic Dataset Distillation
by: Li, Wenyuan, et al.
Published: (2025)
by: Li, Wenyuan, et al.
Published: (2025)
HALC: Object Hallucination Reduction via Adaptive Focal-Contrast Decoding
by: Chen, Zhaorun, et al.
Published: (2024)
by: Chen, Zhaorun, et al.
Published: (2024)
REMEDI: Relative Feature Enhanced Meta-Learning with Distillation for Imbalanced Prediction
by: Liu, Fei, et al.
Published: (2025)
by: Liu, Fei, et al.
Published: (2025)
Large Language Model Guided Knowledge Distillation for Time Series Anomaly Detection
by: Liu, Chen, et al.
Published: (2024)
by: Liu, Chen, et al.
Published: (2024)
Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
by: Wang, Chaoqi, et al.
Published: (2025)
by: Wang, Chaoqi, et al.
Published: (2025)
Distilling Time Series Foundation Models for Efficient Forecasting
by: Li, Yuqi, et al.
Published: (2026)
by: Li, Yuqi, et al.
Published: (2026)
Risk Awareness Injection: Calibrating Vision-Language Models for Safety without Compromising Utility
by: Wang, Mengxuan, et al.
Published: (2026)
by: Wang, Mengxuan, et al.
Published: (2026)
TRACE: Distilling Where It Matters via Token-Routed Self On-Policy Alignment
by: Wang, Jiaxuan, et al.
Published: (2026)
by: Wang, Jiaxuan, et al.
Published: (2026)
Beyond Linear Approximations: A Novel Pruning Approach for Attention Matrix
by: Liang, Yingyu, et al.
Published: (2024)
by: Liang, Yingyu, et al.
Published: (2024)
Looped ReLU MLPs May Be All You Need as Practical Programmable Computers
by: Liang, Yingyu, et al.
Published: (2024)
by: Liang, Yingyu, et al.
Published: (2024)
Multi-Layer Transformers Gradient Can be Approximated in Almost Linear Time
by: Liang, Yingyu, et al.
Published: (2024)
by: Liang, Yingyu, et al.
Published: (2024)
Similar Items
-
Not All Samples Should Be Utilized Equally: Towards Understanding and Improving Dataset Distillation
by: Wang, Shaobo, et al.
Published: (2024) -
DRUPI: Dataset Reduction Using Privileged Information
by: Wang, Shaobo, et al.
Published: (2024) -
Dataset Distillation with Neural Characteristic Function: A Minmax Perspective
by: Wang, Shaobo, et al.
Published: (2025) -
Towards Principled Dataset Distillation: A Spectral Distribution Perspective
by: Wu, Ruixi, et al.
Published: (2026) -
SpeCa: Accelerating Diffusion Transformers with Speculative Feature Caching
by: Liu, Jiacheng, et al.
Published: (2025)