Exploring Learning Complexity for Efficient Downstream Dataset Pruning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Wenyu, Liu, Zhenlong, Xie, Zejian, Zhang, Songxin, Jing, Bingyi, Wei, Hongxin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Fine-tuning can Help Detect Pretraining Data from Large Language Models
von: Zhang, Hengxiang, et al.
Veröffentlicht: (2024)
von: Zhang, Hengxiang, et al.
Veröffentlicht: (2024)
How does Bayesian Sampling help Membership Inference Attacks?
von: Liu, Zhenlong, et al.
Veröffentlicht: (2025)
von: Liu, Zhenlong, et al.
Veröffentlicht: (2025)
Exploring Imbalanced Annotations for Effective In-Context Learning
von: Gao, Hongfu, et al.
Veröffentlicht: (2025)
von: Gao, Hongfu, et al.
Veröffentlicht: (2025)
Provable Joint Decontamination for Benchmarking Multiple Large Language Models
von: Liu, Zhenlong, et al.
Veröffentlicht: (2026)
von: Liu, Zhenlong, et al.
Veröffentlicht: (2026)
RACER: Risk-Aware Calibrated Efficient Routing for Large Language Models
von: Hao, Sai, et al.
Veröffentlicht: (2026)
von: Hao, Sai, et al.
Veröffentlicht: (2026)
Toward Early Quality Assessment of Text-to-Image Diffusion Models
von: Guo, Huanlei, et al.
Veröffentlicht: (2026)
von: Guo, Huanlei, et al.
Veröffentlicht: (2026)
BatchWeave: A Consistent Object-Store-Native Data Plane for Large Foundation Model Training
von: Sun, Ting, et al.
Veröffentlicht: (2026)
von: Sun, Ting, et al.
Veröffentlicht: (2026)
On the Provable Performance Guarantee of Efficient Reasoning Models
von: Zeng, Hao, et al.
Veröffentlicht: (2025)
von: Zeng, Hao, et al.
Veröffentlicht: (2025)
Parametric Scaling Law of Tuning Bias in Conformal Prediction
von: Zeng, Hao, et al.
Veröffentlicht: (2025)
von: Zeng, Hao, et al.
Veröffentlicht: (2025)
Provable Training Data Identification for Large Language Models
von: Liu, Zhenlong, et al.
Veröffentlicht: (2025)
von: Liu, Zhenlong, et al.
Veröffentlicht: (2025)
HyPAC: Cost-Efficient LLMs-Human Hybrid Annotation with PAC Error Guarantees
von: Zeng, Hao, et al.
Veröffentlicht: (2026)
von: Zeng, Hao, et al.
Veröffentlicht: (2026)
Semi-Supervised Conformal Prediction With Unlabeled Nonconformity Score
von: Zhou, Xuanning, et al.
Veröffentlicht: (2025)
von: Zhou, Xuanning, et al.
Veröffentlicht: (2025)
Mitigating Privacy Risk in Membership Inference by Convex-Concave Loss
von: Liu, Zhenlong, et al.
Veröffentlicht: (2024)
von: Liu, Zhenlong, et al.
Veröffentlicht: (2024)
On the Noise Robustness of In-Context Learning for Text Generation
von: Gao, Hongfu, et al.
Veröffentlicht: (2024)
von: Gao, Hongfu, et al.
Veröffentlicht: (2024)
DOS: Diverse Outlier Sampling for Out-of-Distribution Detection
von: Jiang, Wenyu, et al.
Veröffentlicht: (2023)
von: Jiang, Wenyu, et al.
Veröffentlicht: (2023)
TorchCP: A Python Library for Conformal Prediction
von: Huang, Jianguo, et al.
Veröffentlicht: (2024)
von: Huang, Jianguo, et al.
Veröffentlicht: (2024)
Exploring 3D Dataset Pruning
von: Zhao, Xiaohan, et al.
Veröffentlicht: (2026)
von: Zhao, Xiaohan, et al.
Veröffentlicht: (2026)
Learning from Complexity: Exploring Dynamic Sample Pruning of Spatio-Temporal Training
von: Chen, Wei, et al.
Veröffentlicht: (2026)
von: Chen, Wei, et al.
Veröffentlicht: (2026)
Understanding and Mitigating the Label Noise in Pre-training on Downstream Tasks
von: Chen, Hao, et al.
Veröffentlicht: (2023)
von: Chen, Hao, et al.
Veröffentlicht: (2023)
Cut Your Losses! Learning to Prune Paths Early for Efficient Parallel Reasoning
von: Bi, Jiaxi, et al.
Veröffentlicht: (2026)
von: Bi, Jiaxi, et al.
Veröffentlicht: (2026)
Natural Language-Driven Global Mapping of Martian Landforms
von: Wang, Yiran, et al.
Veröffentlicht: (2026)
von: Wang, Yiran, et al.
Veröffentlicht: (2026)
Dataset Representativeness and Downstream Task Fairness
von: Borza, Victor, et al.
Veröffentlicht: (2024)
von: Borza, Victor, et al.
Veröffentlicht: (2024)
Exploring the Noise Robustness of Online Conformal Prediction
von: Xi, Huajun, et al.
Veröffentlicht: (2025)
von: Xi, Huajun, et al.
Veröffentlicht: (2025)
Secure Transfer Learning: Training Clean Models Against Backdoor in (Both) Pre-trained Encoders and Downstream Datasets
von: Zhang, Yechao, et al.
Veröffentlicht: (2025)
von: Zhang, Yechao, et al.
Veröffentlicht: (2025)
CipherPrune: Efficient and Scalable Private Transformer Inference
von: Zhang, Yancheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yancheng, et al.
Veröffentlicht: (2025)
PVI: Plug-in Visual Injection for Vision-Language-Action Models
von: Zhang, Zezhou, et al.
Veröffentlicht: (2026)
von: Zhang, Zezhou, et al.
Veröffentlicht: (2026)
Multimodal-Guided Dynamic Dataset Pruning for Robust and Efficient Data-Centric Learning
von: Yang, Suorong, et al.
Veröffentlicht: (2025)
von: Yang, Suorong, et al.
Veröffentlicht: (2025)
FAIR-Pruner: A Flexible Framework for Automatic Layer-Wise Pruning via Tolerance of Difference
von: Lin, Chenqing, et al.
Veröffentlicht: (2025)
von: Lin, Chenqing, et al.
Veröffentlicht: (2025)
FALCON: FLOP-Aware Combinatorial Optimization for Neural Network Pruning
von: Meng, Xiang, et al.
Veröffentlicht: (2024)
von: Meng, Xiang, et al.
Veröffentlicht: (2024)
Downstream-Pretext Domain Knowledge Traceback for Active Learning
von: Zhang, Beichen, et al.
Veröffentlicht: (2024)
von: Zhang, Beichen, et al.
Veröffentlicht: (2024)
Energy-Efficient Deep Reinforcement Learning with Spiking Transformers
von: Uddin, Mohammad Irfan, et al.
Veröffentlicht: (2025)
von: Uddin, Mohammad Irfan, et al.
Veröffentlicht: (2025)
Lightweight Edge Learning via Dataset Pruning
von: Ale, Laha, et al.
Veröffentlicht: (2026)
von: Ale, Laha, et al.
Veröffentlicht: (2026)
PrunePEFT: Iterative Hybrid Pruning for Parameter-Efficient Fine-tuning of LLMs
von: Yu, Tongzhou, et al.
Veröffentlicht: (2025)
von: Yu, Tongzhou, et al.
Veröffentlicht: (2025)
Exploring Federated Pruning for Large Language Models
von: Guo, Pengxin, et al.
Veröffentlicht: (2025)
von: Guo, Pengxin, et al.
Veröffentlicht: (2025)
Junk DNA Hypothesis: Pruning Small Pre-Trained Weights Irreversibly and Monotonically Impairs "Difficult" Downstream Tasks in LLMs
von: Yin, Lu, et al.
Veröffentlicht: (2023)
von: Yin, Lu, et al.
Veröffentlicht: (2023)
FedBAP: Backdoor Defense via Benign Adversarial Perturbation in Federated Learning
von: Yan, Xinhai, et al.
Veröffentlicht: (2025)
von: Yan, Xinhai, et al.
Veröffentlicht: (2025)
Robust Graph Structure Learning under Heterophily
von: Xie, Xuanting, et al.
Veröffentlicht: (2024)
von: Xie, Xuanting, et al.
Veröffentlicht: (2024)
ST-BCP: Tightening Coverage Bound for Backward Conformal Prediction via Non-Conformity Score Transformation
von: Liu, Junxian, et al.
Veröffentlicht: (2026)
von: Liu, Junxian, et al.
Veröffentlicht: (2026)
SHRP: Specialized Head Routing and Pruning for Efficient Encoder Compression
von: Su, Zeli, et al.
Veröffentlicht: (2025)
von: Su, Zeli, et al.
Veröffentlicht: (2025)
Exploring Vision Neural Network Pruning via Screening Methodology
von: Wang, Mingyuan, et al.
Veröffentlicht: (2025)
von: Wang, Mingyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Fine-tuning can Help Detect Pretraining Data from Large Language Models
von: Zhang, Hengxiang, et al.
Veröffentlicht: (2024) -
How does Bayesian Sampling help Membership Inference Attacks?
von: Liu, Zhenlong, et al.
Veröffentlicht: (2025) -
Exploring Imbalanced Annotations for Effective In-Context Learning
von: Gao, Hongfu, et al.
Veröffentlicht: (2025) -
Provable Joint Decontamination for Benchmarking Multiple Large Language Models
von: Liu, Zhenlong, et al.
Veröffentlicht: (2026) -
RACER: Risk-Aware Calibrated Efficient Routing for Large Language Models
von: Hao, Sai, et al.
Veröffentlicht: (2026)