Utility Boundary of Dataset Distillation: Scaling and Configuration-Coverage Laws
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Zhengquan, Xu, Zhiqiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Path-Guided Flow Matching for Dataset Distillation
by: Li, Xuhui, et al.
Published: (2026)
by: Li, Xuhui, et al.
Published: (2026)
Local-Curvature-Aware Knowledge Graph Embedding: An Extended Ricci Flow Approach
by: Luo, Zhengquan, et al.
Published: (2025)
by: Luo, Zhengquan, et al.
Published: (2025)
KGOT: Unified Knowledge Graph and Optimal Transport Pseudo-Labeling for Molecule-Protein Interaction Prediction
by: Qin, Jiayu, et al.
Published: (2025)
by: Qin, Jiayu, et al.
Published: (2025)
Grounding and Enhancing Informativeness and Utility in Dataset Distillation
by: Wang, Shaobo, et al.
Published: (2026)
by: Wang, Shaobo, et al.
Published: (2026)
Configuration-to-Performance Scaling Law with Neural Ansatz
by: Zhang, Huaqing, et al.
Published: (2026)
by: Zhang, Huaqing, et al.
Published: (2026)
Distillation Scaling Laws
by: Busbridge, Dan, et al.
Published: (2025)
by: Busbridge, Dan, et al.
Published: (2025)
GeoDM: Geometry-aware Distribution Matching for Dataset Distillation
by: Li, Xuhui, et al.
Published: (2025)
by: Li, Xuhui, et al.
Published: (2025)
Tokens-per-Parameter Coverage Is Critical for Robust LLM Scaling Law Extrapolation
by: Kricheli, Joshua Shay, et al.
Published: (2026)
by: Kricheli, Joshua Shay, et al.
Published: (2026)
Dataset Distillation as Data Compression: A Rate-Utility Perspective
by: Bao, Youneng, et al.
Published: (2025)
by: Bao, Youneng, et al.
Published: (2025)
Scaling Laws for Online Advertisement Retrieval
by: Wang, Yunli, et al.
Published: (2024)
by: Wang, Yunli, et al.
Published: (2024)
CK4Gen: A Knowledge Distillation Framework for Generating High-Utility Synthetic Survival Datasets in Healthcare
by: Kuo, Nicholas I-Hsien, et al.
Published: (2024)
by: Kuo, Nicholas I-Hsien, et al.
Published: (2024)
High-dimensional Analysis of Knowledge Distillation: Weak-to-Strong Generalization and Scaling Laws
by: Ildiz, M. Emrullah, et al.
Published: (2024)
by: Ildiz, M. Emrullah, et al.
Published: (2024)
Dataset Distillation via Curriculum Data Synthesis in Large Data Era
by: Yin, Zeyuan, et al.
Published: (2023)
by: Yin, Zeyuan, et al.
Published: (2023)
Using Scaling Laws for Data Source Utility Estimation in Domain-Specific Pre-Training
by: Ostapenko, Oleksiy, et al.
Published: (2025)
by: Ostapenko, Oleksiy, et al.
Published: (2025)
Linux Kernel Configurations at Scale: A Dataset for Performance and Evolution Analysis
by: Borges, Heraldo, et al.
Published: (2025)
by: Borges, Heraldo, et al.
Published: (2025)
Beyond Scaling Law: A Data-Efficient Distillation Framework for Reasoning
by: Wu, Xiaojun, et al.
Published: (2025)
by: Wu, Xiaojun, et al.
Published: (2025)
Privacy-Preserving Federated Learning via Dataset Distillation
by: Xu, ShiMao, et al.
Published: (2024)
by: Xu, ShiMao, et al.
Published: (2024)
FBI-LLM: Scaling Up Fully Binarized LLMs from Scratch via Autoregressive Distillation
by: Ma, Liqun, et al.
Published: (2024)
by: Ma, Liqun, et al.
Published: (2024)
Loss-to-Loss Prediction: Scaling Laws for All Datasets
by: Brandfonbrener, David, et al.
Published: (2024)
by: Brandfonbrener, David, et al.
Published: (2024)
What is Dataset Distillation Learning?
by: Yang, William, et al.
Published: (2024)
by: Yang, William, et al.
Published: (2024)
Distilling Long-tailed Datasets
by: Zhao, Zhenghao, et al.
Published: (2024)
by: Zhao, Zhenghao, et al.
Published: (2024)
Accelerating Large-Scale Dataset Distillation via Exploration-Exploitation Optimization
by: Alahmadi, Muhammad J., et al.
Published: (2026)
by: Alahmadi, Muhammad J., et al.
Published: (2026)
DELT: A Simple Diversity-driven EarlyLate Training for Dataset Distillation
by: Shen, Zhiqiang, et al.
Published: (2024)
by: Shen, Zhiqiang, et al.
Published: (2024)
Towards Trustworthy Dataset Distillation
by: Ma, Shijie, et al.
Published: (2023)
by: Ma, Shijie, et al.
Published: (2023)
Multi-Modal Dataset Distillation in the Wild
by: Dang, Zhuohang, et al.
Published: (2025)
by: Dang, Zhuohang, et al.
Published: (2025)
Diffusion Models as Dataset Distillation Priors
by: Su, Duo, et al.
Published: (2025)
by: Su, Duo, et al.
Published: (2025)
Technical Report on Text Dataset Distillation
by: Ogawa, Keith Ando, et al.
Published: (2025)
by: Ogawa, Keith Ando, et al.
Published: (2025)
Distributional Dataset Distillation with Subtask Decomposition
by: Qin, Tian, et al.
Published: (2024)
by: Qin, Tian, et al.
Published: (2024)
Scaling Laws are Redundancy Laws
by: Bi, Yuda, et al.
Published: (2025)
by: Bi, Yuda, et al.
Published: (2025)
DiLM: Distilling Dataset into Language Model for Text-level Dataset Distillation
by: Maekawa, Aru, et al.
Published: (2024)
by: Maekawa, Aru, et al.
Published: (2024)
Aligning Dense Retrievers with LLM Utility via Distillation
by: Sandhu, Rajinder, et al.
Published: (2026)
by: Sandhu, Rajinder, et al.
Published: (2026)
Towards Stable and Storage-efficient Dataset Distillation: Matching Convexified Trajectory
by: Zhong, Wenliang, et al.
Published: (2024)
by: Zhong, Wenliang, et al.
Published: (2024)
MaskTab: Scalable Masked Tabular Pretraining with Scaling Laws and Distillation for Industrial Classification
by: Zheng, Bo, et al.
Published: (2026)
by: Zheng, Bo, et al.
Published: (2026)
Federated Language Models Under Bandwidth Budgets: Distillation Rates and Conformal Coverage
by: Dubey, Prasanjit, et al.
Published: (2026)
by: Dubey, Prasanjit, et al.
Published: (2026)
Scaling Law for Quantization-Aware Training
by: Chen, Mengzhao, et al.
Published: (2025)
by: Chen, Mengzhao, et al.
Published: (2025)
Distilling Reinforcement Learning into Single-Batch Datasets
by: Wilhelm, Connor, et al.
Published: (2025)
by: Wilhelm, Connor, et al.
Published: (2025)
Finding Stable Subnetworks at Initialization with Dataset Distillation
by: McDermott, Luke, et al.
Published: (2025)
by: McDermott, Luke, et al.
Published: (2025)
Harmonic Dataset Distillation for Time Series Forecasting
by: Hong, Seungha, et al.
Published: (2026)
by: Hong, Seungha, et al.
Published: (2026)
Self-Supervised Dataset Distillation for Transfer Learning
by: Lee, Dong Bok, et al.
Published: (2023)
by: Lee, Dong Bok, et al.
Published: (2023)
Temporal Saliency-Guided Distillation: A Scalable Framework for Distilling Video Datasets
by: Gu, Xulin, et al.
Published: (2025)
by: Gu, Xulin, et al.
Published: (2025)
Similar Items
-
Path-Guided Flow Matching for Dataset Distillation
by: Li, Xuhui, et al.
Published: (2026) -
Local-Curvature-Aware Knowledge Graph Embedding: An Extended Ricci Flow Approach
by: Luo, Zhengquan, et al.
Published: (2025) -
KGOT: Unified Knowledge Graph and Optimal Transport Pseudo-Labeling for Molecule-Protein Interaction Prediction
by: Qin, Jiayu, et al.
Published: (2025) -
Grounding and Enhancing Informativeness and Utility in Dataset Distillation
by: Wang, Shaobo, et al.
Published: (2026) -
Configuration-to-Performance Scaling Law with Neural Ansatz
by: Zhang, Huaqing, et al.
Published: (2026)