Class-Proportional Coreset Selection for Difficulty-Separable Data
Fuente:
arXiv
Saved in:
| Main Authors: | Tsai, Elisa, Zheng, Haizhong, Prakash, Atul |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ELFS: Label-Free Coreset Selection with Proxy Training Dynamics
by: Zheng, Haizhong, et al.
Published: (2024)
by: Zheng, Haizhong, et al.
Published: (2024)
Non-Uniform Class-Wise Coreset Selection for Vision Model Fine-tuning
by: Zhang, Hanyu, et al.
Published: (2025)
by: Zhang, Hanyu, et al.
Published: (2025)
MODE: Multi-Objective Adaptive Coreset Selection
by: Mukherjee, Tanmoy, et al.
Published: (2025)
by: Mukherjee, Tanmoy, et al.
Published: (2025)
Prosperity before Collapse: How Far Can Off-Policy RL Reach with Stale Data on LLMs?
by: Zheng, Haizhong, et al.
Published: (2025)
by: Zheng, Haizhong, et al.
Published: (2025)
Learn To be Efficient: Build Structured Sparsity in Large Language Models
by: Zheng, Haizhong, et al.
Published: (2024)
by: Zheng, Haizhong, et al.
Published: (2024)
Speculative Coreset Selection for Task-Specific Fine-tuning
by: Zhang, Xiaoyu, et al.
Published: (2024)
by: Zhang, Xiaoyu, et al.
Published: (2024)
HyperCore: Coreset Selection under Noise via Hypersphere Models
by: Moser, Brian B., et al.
Published: (2025)
by: Moser, Brian B., et al.
Published: (2025)
FAST: Topology-Aware Frequency-Domain Distribution Matching for Coreset Selection
by: Cui, Jin, et al.
Published: (2025)
by: Cui, Jin, et al.
Published: (2025)
TAGCOS: Task-agnostic Gradient Clustered Coreset Selection for Instruction Tuning Data
by: Zhang, Jipeng, et al.
Published: (2024)
by: Zhang, Jipeng, et al.
Published: (2024)
SubZeroCore: A Submodular Approach with Zero Training for Coreset Selection
by: Moser, Brian B., et al.
Published: (2025)
by: Moser, Brian B., et al.
Published: (2025)
UPCORE: Utility-Preserving Coreset Selection for Balanced Unlearning
by: Patil, Vaidehi, et al.
Published: (2025)
by: Patil, Vaidehi, et al.
Published: (2025)
\emph{FoQuS}: A Forgetting-Quality Coreset Selection Framework for Automatic Modulation Recognition
by: Lu, Yao, et al.
Published: (2025)
by: Lu, Yao, et al.
Published: (2025)
Bayesian Pseudo-Coresets via Contrastive Divergence
by: Tiwary, Piyush, et al.
Published: (2023)
by: Tiwary, Piyush, et al.
Published: (2023)
Spectral Greedy Coresets for Graph Neural Networks
by: Ding, Mucong, et al.
Published: (2024)
by: Ding, Mucong, et al.
Published: (2024)
Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts
by: Zheng, Haizhong, et al.
Published: (2025)
by: Zheng, Haizhong, et al.
Published: (2025)
Finding the Muses: Identifying Coresets through Loss Trajectories
by: Nagaraj, Manish, et al.
Published: (2025)
by: Nagaraj, Manish, et al.
Published: (2025)
Difficulty-Based Preference Data Selection by DPO Implicit Reward Gap
by: Qi, Xuan, et al.
Published: (2025)
by: Qi, Xuan, et al.
Published: (2025)
Predefined Prototypes for Intra-Class Separation and Disentanglement
by: Almudévar, Antonio, et al.
Published: (2024)
by: Almudévar, Antonio, et al.
Published: (2024)
To Predict or Not To Predict? Proportionally Masked Autoencoders for Tabular Data Imputation
by: Kim, Jungkyu, et al.
Published: (2024)
by: Kim, Jungkyu, et al.
Published: (2024)
Mini-batch Coresets for Memory-efficient Language Model Training on Data Mixtures
by: Nguyen, Dang, et al.
Published: (2024)
by: Nguyen, Dang, et al.
Published: (2024)
Efficient and Robust Quantization-aware Training via Adaptive Coreset Selection
by: Huang, Xijie, et al.
Published: (2023)
by: Huang, Xijie, et al.
Published: (2023)
On the Difficulty of Learning a Meta-network for Training Data Selection
by: Du, Zilin, et al.
Published: (2026)
by: Du, Zilin, et al.
Published: (2026)
Data Difficulty and the Generalization--Extrapolation Tradeoff in LLM Fine-Tuning
by: Liu, Siyuan, et al.
Published: (2026)
by: Liu, Siyuan, et al.
Published: (2026)
Noise-Tolerant Coreset-Based Class Incremental Continual Learning
by: Mucllari, Edison, et al.
Published: (2025)
by: Mucllari, Edison, et al.
Published: (2025)
GraphSculptor: Sculpting Pre-training Coreset for Graph Self-supervised Learning
by: Liu, Chuang, et al.
Published: (2026)
by: Liu, Chuang, et al.
Published: (2026)
DARE: Difficulty-Adaptive Reinforcement Learning with Co-Evolved Difficulty Estimation
by: Zhou, Yang, et al.
Published: (2026)
by: Zhou, Yang, et al.
Published: (2026)
3DS: Medical Domain Adaptation of LLMs via Decomposed Difficulty-based Data Selection
by: Ding, Hongxin, et al.
Published: (2024)
by: Ding, Hongxin, et al.
Published: (2024)
Improving Data Efficiency for LLM Reinforcement Fine-tuning Through Difficulty-targeted Online Data Selection and Rollout Replay
by: Sun, Yifan, et al.
Published: (2025)
by: Sun, Yifan, et al.
Published: (2025)
Utilizing Class Separation Distance for the Evaluation of Corruption Robustness of Machine Learning Classifiers
by: Siedel, Georg, et al.
Published: (2022)
by: Siedel, Georg, et al.
Published: (2022)
Preference Models assume Proportional Hazards of Utilities
by: Nagpal, Chirag
Published: (2025)
by: Nagpal, Chirag
Published: (2025)
Optimistic Rates for Learning from Label Proportions
by: Li, Gene, et al.
Published: (2024)
by: Li, Gene, et al.
Published: (2024)
Uncertainty-Aware Deep Neural Representations for Visual Analysis of Vector Field Data
by: Kumar, Atul, et al.
Published: (2024)
by: Kumar, Atul, et al.
Published: (2024)
Taming Data Challenges in ML-based Security Tasks Using Generative AI
by: Kanchi, Shravya, et al.
Published: (2025)
by: Kanchi, Shravya, et al.
Published: (2025)
MathMixup: Boosting LLM Mathematical Reasoning with Difficulty-Controllable Data Synthesis and Curriculum Learning
by: Li, Xuchen, et al.
Published: (2026)
by: Li, Xuchen, et al.
Published: (2026)
Diversity-Aware Adaptive Collocation for Physics-Informed Neural Networks via Sparse QUBO Optimization and Hybrid Coresets
by: Salloum, Hadi, et al.
Published: (2026)
by: Salloum, Hadi, et al.
Published: (2026)
Transparency and Proportionality in Post-Processing Algorithmic Bias Correction
by: Ferreira, Juliett Suárez, et al.
Published: (2025)
by: Ferreira, Juliett Suárez, et al.
Published: (2025)
Evaluating Game Difficulty in Tetris Block Puzzle
by: Wang, Chun-Jui, et al.
Published: (2026)
by: Wang, Chun-Jui, et al.
Published: (2026)
Optimizing Reasoning Efficiency through Prompt Difficulty Prediction
by: Zhao, Bo, et al.
Published: (2025)
by: Zhao, Bo, et al.
Published: (2025)
LLM Unlearning Reveals a Stronger-Than-Expected Coreset Effect in Current Benchmarks
by: Pal, Soumyadeep, et al.
Published: (2025)
by: Pal, Soumyadeep, et al.
Published: (2025)
Beyond RLHF and NLHF: Population-Proportional Alignment under an Axiomatic Framework
by: Kim, Kihyun, et al.
Published: (2025)
by: Kim, Kihyun, et al.
Published: (2025)
Similar Items
-
ELFS: Label-Free Coreset Selection with Proxy Training Dynamics
by: Zheng, Haizhong, et al.
Published: (2024) -
Non-Uniform Class-Wise Coreset Selection for Vision Model Fine-tuning
by: Zhang, Hanyu, et al.
Published: (2025) -
MODE: Multi-Objective Adaptive Coreset Selection
by: Mukherjee, Tanmoy, et al.
Published: (2025) -
Prosperity before Collapse: How Far Can Off-Policy RL Reach with Stale Data on LLMs?
by: Zheng, Haizhong, et al.
Published: (2025) -
Learn To be Efficient: Build Structured Sparsity in Large Language Models
by: Zheng, Haizhong, et al.
Published: (2024)