Taming Hyperparameter Sensitivity in Data Attribution: Practical Selection Without Costly Retraining
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Weiyi, Deng, Junwei, Hu, Yuzheng, Zhang, Shiyuan, Jiang, Xirui, Zhang, Runting, Zhao, Han, Ma, Jiaqi W. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring Training Data Attribution under Limited Access Constraints
by: Zhang, Shiyuan, et al.
Published: (2025)
by: Zhang, Shiyuan, et al.
Published: (2025)
Adversarial Attacks on Data Attribution
by: Wang, Xinhe, et al.
Published: (2024)
by: Wang, Xinhe, et al.
Published: (2024)
Computational Copyright: Towards A Royalty Model for Music Generative AI
by: Deng, Junwei, et al.
Published: (2023)
by: Deng, Junwei, et al.
Published: (2023)
Most Influential Subset Selection: Challenges, Promises, and Beyond
by: Hu, Yuzheng, et al.
Published: (2024)
by: Hu, Yuzheng, et al.
Published: (2024)
How Faithful Is Trajectory-Based Data Attribution? Error Sources, Remedies, and Practical Guidelines
by: Deng, Junwei, et al.
Published: (2026)
by: Deng, Junwei, et al.
Published: (2026)
A Versatile Influence Function for Data Attribution with Non-Decomposable Loss
by: Deng, Junwei, et al.
Published: (2024)
by: Deng, Junwei, et al.
Published: (2024)
$\texttt{dattri}$: A Library for Efficient Data Attribution
by: Deng, Junwei, et al.
Published: (2024)
by: Deng, Junwei, et al.
Published: (2024)
Efficient Ensembles Improve Training Data Attribution
by: Deng, Junwei, et al.
Published: (2024)
by: Deng, Junwei, et al.
Published: (2024)
A Snapshot of Influence: A Local Data Attribution Framework for Online Reinforcement Learning
by: Hu, Yuzheng, et al.
Published: (2025)
by: Hu, Yuzheng, et al.
Published: (2025)
A Unified Theory of Random Projection for Influence Functions
by: Hu, Pingbang, et al.
Published: (2026)
by: Hu, Pingbang, et al.
Published: (2026)
GraSS: Scalable Data Attribution with Gradient Sparsification and Sparse Projection
by: Hu, Pingbang, et al.
Published: (2025)
by: Hu, Pingbang, et al.
Published: (2025)
From Stealthy Data Fabrication to Unsafe Driving: Realistic Scenario Attacks on Collaborative Perception
by: Zhang, Qingzhao, et al.
Published: (2026)
by: Zhang, Qingzhao, et al.
Published: (2026)
Cost-Sensitive Freeze-thaw Bayesian Optimization for Efficient Hyperparameter Tuning
by: Lee, Dong Bok, et al.
Published: (2025)
by: Lee, Dong Bok, et al.
Published: (2025)
Detecting and Filtering Unsafe Training Data via Data Attribution with Denoised Representation
by: Pan, Yijun, et al.
Published: (2025)
by: Pan, Yijun, et al.
Published: (2025)
Daunce: Data Attribution through Uncertainty Estimation
by: Pan, Xingyuan, et al.
Published: (2025)
by: Pan, Xingyuan, et al.
Published: (2025)
OATS: Online Data Augmentation for Time Series Foundation Models
by: Deng, Junwei, et al.
Published: (2026)
by: Deng, Junwei, et al.
Published: (2026)
Merging Models on the Fly Without Retraining: A Sequential Approach to Scalable Continual Model Merging
by: Tang, Anke, et al.
Published: (2025)
by: Tang, Anke, et al.
Published: (2025)
Conditional stability in determining source terms of semilinear parabolic partial differential equations
by: Xirui, Hu
Published: (2025)
by: Xirui, Hu
Published: (2025)
Taming Sampling Perturbations with Variance Expansion Loss for Latent Diffusion Models
by: Li, Qifan, et al.
Published: (2026)
by: Li, Qifan, et al.
Published: (2026)
DiP: Taming Diffusion Models in Pixel Space
by: Chen, Zhennan, et al.
Published: (2025)
by: Chen, Zhennan, et al.
Published: (2025)
Prediction of Coffee Ratings Based On Influential Attributes Using SelectKBest and Optimal Hyperparameters
by: Agyemang, Edmund, et al.
Published: (2025)
by: Agyemang, Edmund, et al.
Published: (2025)
Promoting Unsupervised Data‐To‐Text Generation Using Retraining and Unified Linearization
by: Xiaobo Wang, et al.
Published: (2025)
by: Xiaobo Wang, et al.
Published: (2025)
Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining
by: Abro, Aarash, et al.
Published: (2026)
by: Abro, Aarash, et al.
Published: (2026)
Solvent‐Free, High Castor Oil Content‐Based Pressure‐Sensitive Adhesive: Synthesis and Performance
by: Yangbao Ma, et al.
Published: (2025)
by: Yangbao Ma, et al.
Published: (2025)
Measuring Fine-Grained Relatedness in Multitask Learning via Data Attribution
by: Tu, Yiwen, et al.
Published: (2025)
by: Tu, Yiwen, et al.
Published: (2025)
Localize-and-Stitch: Efficient Model Merging via Sparse Task Arithmetic
by: He, Yifei, et al.
Published: (2024)
by: He, Yifei, et al.
Published: (2024)
Greedy Output Approximation: Towards Efficient Structured Pruning for LLMs Without Retraining
by: Li, Jianwei, et al.
Published: (2024)
by: Li, Jianwei, et al.
Published: (2024)
Retrieval Augmented Anomaly Detection (RAAD): Nimble Model Adjustment Without Retraining
by: Pastoriza, Sam, et al.
Published: (2025)
by: Pastoriza, Sam, et al.
Published: (2025)
Sustainable Machine Learning Retraining: Optimizing Energy Efficiency Without Compromising Accuracy
by: Poenaru-Olaru, Lorena, et al.
Published: (2025)
by: Poenaru-Olaru, Lorena, et al.
Published: (2025)
Leveraging Computational Pathology AI for Noninvasive Optical Imaging Analysis Without Retraining
by: Barash, Danny, et al.
Published: (2024)
by: Barash, Danny, et al.
Published: (2024)
Who Gets Credit or Blame? Attributing Accountability in Modern AI Systems
by: Zhang, Shichang, et al.
Published: (2025)
by: Zhang, Shichang, et al.
Published: (2025)
Challenge to Eliminate Parasite Palatal Infestation in a Deaf–Mute Geriatric Patient
by: Amelia Elizabeth Pranoto, et al.
Published: (2025)
by: Amelia Elizabeth Pranoto, et al.
Published: (2025)
Towards Fair Large Language Model-based Recommender Systems without Costly Retraining
by: Li, Jin, et al.
Published: (2026)
by: Li, Jin, et al.
Published: (2026)
Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities
by: Dai, Qirun, et al.
Published: (2025)
by: Dai, Qirun, et al.
Published: (2025)
Pick-and-place Manipulation Across Grippers Without Retraining: A Learning-optimization Diffusion Policy Approach
by: Yao, Xiangtong, et al.
Published: (2025)
by: Yao, Xiangtong, et al.
Published: (2025)
Taming Transformer for Emotion-Controllable Talking Face Generation
by: Zhang, Ziqi, et al.
Published: (2025)
by: Zhang, Ziqi, et al.
Published: (2025)
Hyperparameter Selection in Continual Learning
by: Lee, Thomas L., et al.
Published: (2024)
by: Lee, Thomas L., et al.
Published: (2024)
Selecting Hyperparameters for Tree-Boosting
by: Koster, Floris Jan, et al.
Published: (2026)
by: Koster, Floris Jan, et al.
Published: (2026)
Opportunistic Expert Activation: Batch-Aware Expert Routing for Faster Decode Without Retraining
by: Oncescu, Costin-Andrei, et al.
Published: (2025)
by: Oncescu, Costin-Andrei, et al.
Published: (2025)
QLIP: A Dynamic Quadtree Vision Prior Enhances MLLM Performance Without Retraining
by: Chickering, Kyle R., et al.
Published: (2025)
by: Chickering, Kyle R., et al.
Published: (2025)
Similar Items
-
Exploring Training Data Attribution under Limited Access Constraints
by: Zhang, Shiyuan, et al.
Published: (2025) -
Adversarial Attacks on Data Attribution
by: Wang, Xinhe, et al.
Published: (2024) -
Computational Copyright: Towards A Royalty Model for Music Generative AI
by: Deng, Junwei, et al.
Published: (2023) -
Most Influential Subset Selection: Challenges, Promises, and Beyond
by: Hu, Yuzheng, et al.
Published: (2024) -
How Faithful Is Trajectory-Based Data Attribution? Error Sources, Remedies, and Practical Guidelines
by: Deng, Junwei, et al.
Published: (2026)