In-Run Data Shapley for Adam Optimizer
Fuente:
arXiv
Salvato in:
| Autori principali: | Ding, Meng, Zhang, Zeqing, Wang, Di, Hu, Lijie |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Evaluating Data Influence in Meta Learning
di: Ren, Chenyang, et al.
Pubblicazione: (2025)
di: Ren, Chenyang, et al.
Pubblicazione: (2025)
Understanding the Dynamics of Demonstration Conflict in In-Context Learning
di: Jiao, Difan, et al.
Pubblicazione: (2026)
di: Jiao, Difan, et al.
Pubblicazione: (2026)
PAHQ: Accelerating Automated Circuit Discovery through Mixed-Precision Inference Optimization
di: Wang, Xinhai, et al.
Pubblicazione: (2025)
di: Wang, Xinhai, et al.
Pubblicazione: (2025)
Losing is for Cherishing: Data Valuation Based on Machine Unlearning and Shapley Value
di: Ma, Le, et al.
Pubblicazione: (2025)
di: Ma, Le, et al.
Pubblicazione: (2025)
Fast-DataShapley: Neural Modeling for Training Data Valuation
di: Sun, Haifeng, et al.
Pubblicazione: (2025)
di: Sun, Haifeng, et al.
Pubblicazione: (2025)
Global Evolutionary Steering: Refining Activation Steering Control via Cross-Layer Consistency
di: Jiang, Xinyan, et al.
Pubblicazione: (2026)
di: Jiang, Xinyan, et al.
Pubblicazione: (2026)
Beyond First-Order: Training LLMs with Stochastic Conjugate Subgradients and AdamW
di: Zhang, Di, et al.
Pubblicazione: (2025)
di: Zhang, Di, et al.
Pubblicazione: (2025)
Is Data Shapley Not Better than Random in Data Selection? Ask NASH
di: Tian, Xiao, et al.
Pubblicazione: (2026)
di: Tian, Xiao, et al.
Pubblicazione: (2026)
HyperSHAP: Shapley Values and Interactions for Explaining Hyperparameter Optimization
di: Wever, Marcel, et al.
Pubblicazione: (2025)
di: Wever, Marcel, et al.
Pubblicazione: (2025)
Thresholding Data Shapley for Data Cleansing Using Multi-Armed Bandits
di: Namba, Hiroyuki, et al.
Pubblicazione: (2024)
di: Namba, Hiroyuki, et al.
Pubblicazione: (2024)
KernelSHAP-IQ: Weighted Least-Square Optimization for Shapley Interactions
di: Fumagalli, Fabian, et al.
Pubblicazione: (2024)
di: Fumagalli, Fabian, et al.
Pubblicazione: (2024)
Generalized Priority-Aware Shapley Value
di: Lee, Kiljae, et al.
Pubblicazione: (2026)
di: Lee, Kiljae, et al.
Pubblicazione: (2026)
Why Transformers Need Adam: A Hessian Perspective
di: Zhang, Yushun, et al.
Pubblicazione: (2024)
di: Zhang, Yushun, et al.
Pubblicazione: (2024)
Controlling Repetition in Protein Language Models
di: Zhang, Jiahao, et al.
Pubblicazione: (2026)
di: Zhang, Jiahao, et al.
Pubblicazione: (2026)
Understanding the Impact of Differentially Private Training on Memorization of Long-Tailed Data
di: Zhang, Jiaming, et al.
Pubblicazione: (2026)
di: Zhang, Jiaming, et al.
Pubblicazione: (2026)
An Odd Estimator for Shapley Values
di: Fumagalli, Fabian, et al.
Pubblicazione: (2026)
di: Fumagalli, Fabian, et al.
Pubblicazione: (2026)
On the Inflation of KNN-Shapley Value
di: Yang, Ziao, et al.
Pubblicazione: (2024)
di: Yang, Ziao, et al.
Pubblicazione: (2024)
Suboptimal Shapley Value Explanations
di: Lu, Xiaolei
Pubblicazione: (2025)
di: Lu, Xiaolei
Pubblicazione: (2025)
Helpful or Harmful Data? Fine-tuning-free Shapley Attribution for Explaining Language Model Predictions
di: Wang, Jingtan, et al.
Pubblicazione: (2024)
di: Wang, Jingtan, et al.
Pubblicazione: (2024)
Faithful Interpretation for Graph Neural Networks
di: Hu, Lijie, et al.
Pubblicazione: (2024)
di: Hu, Lijie, et al.
Pubblicazione: (2024)
Optimizer-Induced Mode Connectivity: From AdamW to Muon
di: Zhang, Fangzhao, et al.
Pubblicazione: (2026)
di: Zhang, Fangzhao, et al.
Pubblicazione: (2026)
Adam-mini: Use Fewer Learning Rates To Gain More
di: Zhang, Yushun, et al.
Pubblicazione: (2024)
di: Zhang, Yushun, et al.
Pubblicazione: (2024)
EAP-GP: Mitigating Saturation Effect in Gradient-based Automated Circuit Identification
di: Zhang, Lin, et al.
Pubblicazione: (2025)
di: Zhang, Lin, et al.
Pubblicazione: (2025)
Exactly Computing do-Shapley Values
di: Witter, R. Teal, et al.
Pubblicazione: (2026)
di: Witter, R. Teal, et al.
Pubblicazione: (2026)
Explaining Drift using Shapley Values
di: Edakunni, Narayanan U., et al.
Pubblicazione: (2024)
di: Edakunni, Narayanan U., et al.
Pubblicazione: (2024)
shapiq: Shapley Interactions for Machine Learning
di: Muschalik, Maximilian, et al.
Pubblicazione: (2024)
di: Muschalik, Maximilian, et al.
Pubblicazione: (2024)
Benign Overfitting in Adversarial Training for Vision Transformers
di: Zhang, Jiaming, et al.
Pubblicazione: (2026)
di: Zhang, Jiaming, et al.
Pubblicazione: (2026)
A Multi-Modal CNN-LSTM Framework with Multi-Head Attention and Focal Loss for Real-Time Elderly Fall Detection
di: Zhou, Lijie, et al.
Pubblicazione: (2026)
di: Zhou, Lijie, et al.
Pubblicazione: (2026)
Anon: Extrapolating Adaptivity Beyond SGD and Adam
di: Zhang, Yiheng, et al.
Pubblicazione: (2026)
di: Zhang, Yiheng, et al.
Pubblicazione: (2026)
Shapley-PC: Constraint-based Causal Structure Learning with a Shapley Inspired Framework
di: Russo, Fabrizio, et al.
Pubblicazione: (2023)
di: Russo, Fabrizio, et al.
Pubblicazione: (2023)
The Epochal Sawtooth Phenomenon: Unveiling Training Loss Oscillations in Adam and Other Optimizers
di: Liu, Qi, et al.
Pubblicazione: (2024)
di: Liu, Qi, et al.
Pubblicazione: (2024)
Private Language Models via Truncated Laplacian Mechanism
di: Huang, Tianhao, et al.
Pubblicazione: (2024)
di: Huang, Tianhao, et al.
Pubblicazione: (2024)
Shapley Neuron Values for Continual Learning: Which Neurons Matter Most?
di: Vahedifar, Mohammad Ali, et al.
Pubblicazione: (2026)
di: Vahedifar, Mohammad Ali, et al.
Pubblicazione: (2026)
GRASP: group-Shapley feature selection for patients
di: Luo, Yuheng, et al.
Pubblicazione: (2026)
di: Luo, Yuheng, et al.
Pubblicazione: (2026)
Proxy-Based Approximation of Shapley and Banzhaf Interactions
di: Thies, Santo M. A. R., et al.
Pubblicazione: (2026)
di: Thies, Santo M. A. R., et al.
Pubblicazione: (2026)
Redefining Contributions: Shapley-Driven Federated Learning
di: Tastan, Nurbek, et al.
Pubblicazione: (2024)
di: Tastan, Nurbek, et al.
Pubblicazione: (2024)
Locate-then-edit for Multi-hop Factual Recall under Knowledge Editing
di: Zhang, Zhuoran, et al.
Pubblicazione: (2024)
di: Zhang, Zhuoran, et al.
Pubblicazione: (2024)
AdamS: Momentum Itself Can Be A Normalizer for LLM Pretraining and Post-training
di: Zhang, Huishuai, et al.
Pubblicazione: (2025)
di: Zhang, Huishuai, et al.
Pubblicazione: (2025)
DP-FedAdamW: An Efficient Optimizer for Differentially Private Federated Large Models
di: Liu, Jin, et al.
Pubblicazione: (2026)
di: Liu, Jin, et al.
Pubblicazione: (2026)
Navigating LLM Valley: From AdamW to Memory-Efficient and Matrix-Based Optimizers
di: Ranganath, Aditya
Pubblicazione: (2026)
di: Ranganath, Aditya
Pubblicazione: (2026)
Documenti analoghi
-
Evaluating Data Influence in Meta Learning
di: Ren, Chenyang, et al.
Pubblicazione: (2025) -
Understanding the Dynamics of Demonstration Conflict in In-Context Learning
di: Jiao, Difan, et al.
Pubblicazione: (2026) -
PAHQ: Accelerating Automated Circuit Discovery through Mixed-Precision Inference Optimization
di: Wang, Xinhai, et al.
Pubblicazione: (2025) -
Losing is for Cherishing: Data Valuation Based on Machine Unlearning and Shapley Value
di: Ma, Le, et al.
Pubblicazione: (2025) -
Fast-DataShapley: Neural Modeling for Training Data Valuation
di: Sun, Haifeng, et al.
Pubblicazione: (2025)