Lightweight Dataset Pruning without Full Training via Example Difficulty and Prediction Uncertainty
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cho, Yeseul, Shin, Baekrok, Kang, Changmin, Yun, Chulhee |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DASH: Warm-Starting Neural Network Training in Stationary Settings without Loss of Plasticity
von: Shin, Baekrok, et al.
Veröffentlicht: (2024)
von: Shin, Baekrok, et al.
Veröffentlicht: (2024)
Uniform Spectral Growth and Convergence of Muon in LoRA-Style Matrix Factorization
von: Kang, Changmin, et al.
Veröffentlicht: (2026)
von: Kang, Changmin, et al.
Veröffentlicht: (2026)
Implicit Bias and Loss of Plasticity in Matrix Completion: Depth Promotes Low-Rankness
von: Shin, Baekrok, et al.
Veröffentlicht: (2026)
von: Shin, Baekrok, et al.
Veröffentlicht: (2026)
Understanding Sharpness Dynamics in NN Training with a Minimalist Example: The Effects of Dataset Difficulty, Depth, Stochasticity, and More
von: Yoo, Geonhui, et al.
Veröffentlicht: (2025)
von: Yoo, Geonhui, et al.
Veröffentlicht: (2025)
Lightweight Edge Learning via Dataset Pruning
von: Ale, Laha, et al.
Veröffentlicht: (2026)
von: Ale, Laha, et al.
Veröffentlicht: (2026)
Implicit Bias of Per-sample Adam on Separable Data: Departure from the Full-batch Regime
von: Baek, Beomhan, et al.
Veröffentlicht: (2025)
von: Baek, Beomhan, et al.
Veröffentlicht: (2025)
Provable Benefit of Cutout and CutMix for Feature Learning
von: Oh, Junsoo, et al.
Veröffentlicht: (2024)
von: Oh, Junsoo, et al.
Veröffentlicht: (2024)
Arithmetic Transformers Can Length-Generalize in Both Operand Length and Count
von: Cho, Hanseul, et al.
Veröffentlicht: (2024)
von: Cho, Hanseul, et al.
Veröffentlicht: (2024)
The Cost of Robustness: Tighter Bounds on Parameter Complexity for Robust Memorization in ReLU Nets
von: Kim, Yujun, et al.
Veröffentlicht: (2025)
von: Kim, Yujun, et al.
Veröffentlicht: (2025)
Minor First, Major Last: A Depth-Induced Implicit Bias of Sharpness-Aware Minimization
von: Moon, Chaewon, et al.
Veröffentlicht: (2026)
von: Moon, Chaewon, et al.
Veröffentlicht: (2026)
Through the River: Understanding the Benefit of Schedule-Free Methods for Language Model Training
von: Song, Minhak, et al.
Veröffentlicht: (2025)
von: Song, Minhak, et al.
Veröffentlicht: (2025)
Lightweight and Post-Training Structured Pruning for On-Device Large Lanaguage Models
von: Xu, Zihuai, et al.
Veröffentlicht: (2025)
von: Xu, Zihuai, et al.
Veröffentlicht: (2025)
Scaling Laws of SignSGD in Linear Regression: When Does It Outperform SGD?
von: Kim, Jihwan, et al.
Veröffentlicht: (2026)
von: Kim, Jihwan, et al.
Veröffentlicht: (2026)
Position Coupling: Improving Length Generalization of Arithmetic Transformers Using Task Structure
von: Cho, Hanseul, et al.
Veröffentlicht: (2024)
von: Cho, Hanseul, et al.
Veröffentlicht: (2024)
Parameter Expanded Stochastic Gradient Markov Chain Monte Carlo
von: Kim, Hyunsu, et al.
Veröffentlicht: (2025)
von: Kim, Hyunsu, et al.
Veröffentlicht: (2025)
Can Prompt Difficulty be Online Predicted for Accelerating RL Finetuning of Reasoning Models?
von: Qu, Yun, et al.
Veröffentlicht: (2025)
von: Qu, Yun, et al.
Veröffentlicht: (2025)
Lightweight Geometric Adaptation for Training Physics-Informed Neural Networks
von: An, Kang, et al.
Veröffentlicht: (2026)
von: An, Kang, et al.
Veröffentlicht: (2026)
ECO: Quantized Training without Full-Precision Master Weights
von: Nikdan, Mahdi, et al.
Veröffentlicht: (2026)
von: Nikdan, Mahdi, et al.
Veröffentlicht: (2026)
Training-free LLM Verification via Recycling Few-shot Examples
von: Lee, Dongseok, et al.
Veröffentlicht: (2025)
von: Lee, Dongseok, et al.
Veröffentlicht: (2025)
Zeroth-Order Adaptive Neuron Alignment Based Pruning without Re-Training
von: Cunegatti, Elia, et al.
Veröffentlicht: (2024)
von: Cunegatti, Elia, et al.
Veröffentlicht: (2024)
Influence Functions for Preference Dataset Pruning
von: Fein, Daniel, et al.
Veröffentlicht: (2025)
von: Fein, Daniel, et al.
Veröffentlicht: (2025)
Toward Data Efficient Model Merging between Different Datasets without Performance Degradation
von: Yamada, Masanori, et al.
Veröffentlicht: (2023)
von: Yamada, Masanori, et al.
Veröffentlicht: (2023)
Overcoming Reward Overoptimization via Adversarial Policy Optimization with Lightweight Uncertainty Estimation
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2024)
Optimizing Reasoning Efficiency through Prompt Difficulty Prediction
von: Zhao, Bo, et al.
Veröffentlicht: (2025)
von: Zhao, Bo, et al.
Veröffentlicht: (2025)
Realizing Unaligned Block-wise Pruning for DNN Acceleration on Mobile Devices
von: Lee, Hayun, et al.
Veröffentlicht: (2024)
von: Lee, Hayun, et al.
Veröffentlicht: (2024)
Lightweight Safety Classification Using Pruned Language Models
von: Sawtell, Mason, et al.
Veröffentlicht: (2024)
von: Sawtell, Mason, et al.
Veröffentlicht: (2024)
Pruning and Distilling Mixture-of-Experts into Dense Language Models
von: Kim, Junhyuck, et al.
Veröffentlicht: (2026)
von: Kim, Junhyuck, et al.
Veröffentlicht: (2026)
DONOD: Efficient and Generalizable Instruction Fine-Tuning for LLMs via Model-Intrinsic Dataset Pruning
von: Hu, Jucheng, et al.
Veröffentlicht: (2025)
von: Hu, Jucheng, et al.
Veröffentlicht: (2025)
Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information
von: Ethayarajh, Kawin, et al.
Veröffentlicht: (2021)
von: Ethayarajh, Kawin, et al.
Veröffentlicht: (2021)
Predicting Performance of Symbolic and Prompt Programs with Examples
von: Zheng, Chengqi, et al.
Veröffentlicht: (2026)
von: Zheng, Chengqi, et al.
Veröffentlicht: (2026)
Linear attention is (maybe) all you need (to understand transformer optimization)
von: Ahn, Kwangjun, et al.
Veröffentlicht: (2023)
von: Ahn, Kwangjun, et al.
Veröffentlicht: (2023)
VarDrop: Enhancing Training Efficiency by Reducing Variate Redundancy in Periodic Time Series Forecasting
von: Kang, Junhyeok, et al.
Veröffentlicht: (2025)
von: Kang, Junhyeok, et al.
Veröffentlicht: (2025)
DARE: Difficulty-Adaptive Reinforcement Learning with Co-Evolved Difficulty Estimation
von: Zhou, Yang, et al.
Veröffentlicht: (2026)
von: Zhou, Yang, et al.
Veröffentlicht: (2026)
SwiftPrune: Hessian-Free Weight Pruning for Large Language Models
von: Kang, Yuhan, et al.
Veröffentlicht: (2025)
von: Kang, Yuhan, et al.
Veröffentlicht: (2025)
MemLoss: Enhancing Adversarial Training with Recycling Adversarial Examples
von: Mahdi, Soroush, et al.
Veröffentlicht: (2025)
von: Mahdi, Soroush, et al.
Veröffentlicht: (2025)
Collaborative Learning-Enhanced Lightweight Models for Predicting Arterial Blood Pressure Waveform in a Large-scale Perioperative Dataset
von: Li, Wentao, et al.
Veröffentlicht: (2025)
von: Li, Wentao, et al.
Veröffentlicht: (2025)
TopoPrune: Robust Data Pruning via Unified Latent Space Topology
von: Roy, Arjun, et al.
Veröffentlicht: (2026)
von: Roy, Arjun, et al.
Veröffentlicht: (2026)
Pruning Foundation Models for High Accuracy without Retraining
von: Zhao, Pu, et al.
Veröffentlicht: (2024)
von: Zhao, Pu, et al.
Veröffentlicht: (2024)
Weight Concentration Regularization for Improving Pruning Robustness Under High Sparsity
von: Yun, Vincent-Daniel, et al.
Veröffentlicht: (2025)
von: Yun, Vincent-Daniel, et al.
Veröffentlicht: (2025)
Evaluating Time-Series Training Dataset through Lens of Spectrum in Deep State Space Models
von: Kanai, Sekitoshi, et al.
Veröffentlicht: (2024)
von: Kanai, Sekitoshi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DASH: Warm-Starting Neural Network Training in Stationary Settings without Loss of Plasticity
von: Shin, Baekrok, et al.
Veröffentlicht: (2024) -
Uniform Spectral Growth and Convergence of Muon in LoRA-Style Matrix Factorization
von: Kang, Changmin, et al.
Veröffentlicht: (2026) -
Implicit Bias and Loss of Plasticity in Matrix Completion: Depth Promotes Low-Rankness
von: Shin, Baekrok, et al.
Veröffentlicht: (2026) -
Understanding Sharpness Dynamics in NN Training with a Minimalist Example: The Effects of Dataset Difficulty, Depth, Stochasticity, and More
von: Yoo, Geonhui, et al.
Veröffentlicht: (2025) -
Lightweight Edge Learning via Dataset Pruning
von: Ale, Laha, et al.
Veröffentlicht: (2026)