Towards Robust Influence Functions with Flat Validation Minima
Fuente:
arXiv
Saved in:
| Main Authors: | Ye, Xichen, Wu, Yifan, Zhang, Weizhong, Jin, Cheng, Chen, Yifan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimized Gradient Clipping for Noisy Label Learning
by: Ye, Xichen, et al.
Published: (2024)
by: Ye, Xichen, et al.
Published: (2024)
Towards the Connection between Activation Sparsity and Flat Minima
by: Peng, Ze, et al.
Published: (2026)
by: Peng, Ze, et al.
Published: (2026)
Are Flat Minima an Illusion?
by: Bennett, Michael Timothy
Published: (2026)
by: Bennett, Michael Timothy
Published: (2026)
A Flat Minima Perspective on Understanding Augmentations and Model Robustness
by: Yoo, Weebum, et al.
Published: (2025)
by: Yoo, Weebum, et al.
Published: (2025)
Mirror Gradient: Towards Robust Multimodal Recommender Systems via Exploring Flat Local Minima
by: Zhong, Shanshan, et al.
Published: (2024)
by: Zhong, Shanshan, et al.
Published: (2024)
Active Negative Loss: A Robust Framework for Learning with Noisy Labels
by: Ye, Xichen, et al.
Published: (2024)
by: Ye, Xichen, et al.
Published: (2024)
A Function-Centric Perspective on Flat and Sharp Minima
by: Mason-Williams, Israel, et al.
Published: (2025)
by: Mason-Williams, Israel, et al.
Published: (2025)
Zeroth-Order Optimization Finds Flat Minima
by: Zhang, Liang, et al.
Published: (2025)
by: Zhang, Liang, et al.
Published: (2025)
Embedding Empirical Distributions for Computing Optimal Transport Maps
by: Jiang, Mingchen, et al.
Published: (2025)
by: Jiang, Mingchen, et al.
Published: (2025)
Flat Minima and Generalization: Insights from Stochastic Convex Optimization
by: Schliserman, Matan, et al.
Published: (2025)
by: Schliserman, Matan, et al.
Published: (2025)
A PAC-Bayesian Link Between Generalisation and Flat Minima
by: Haddouche, Maxime, et al.
Published: (2024)
by: Haddouche, Maxime, et al.
Published: (2024)
SAFE: Finding Sparse and Flat Minima to Improve Pruning
by: Lee, Dongyeop, et al.
Published: (2025)
by: Lee, Dongyeop, et al.
Published: (2025)
Computational Budget Should Be Considered in Data Selection
by: Wan, Weilin, et al.
Published: (2025)
by: Wan, Weilin, et al.
Published: (2025)
Seeking Flat Minima over Diverse Surrogates for Improved Adversarial Transferability: A Theoretical Framework and Algorithmic Instantiation
by: Zheng, Meixi, et al.
Published: (2025)
by: Zheng, Meixi, et al.
Published: (2025)
Investigating Data Pruning for Pretraining Biological Foundation Models at Scale
by: Wu, Yifan, et al.
Published: (2025)
by: Wu, Yifan, et al.
Published: (2025)
Holistic Scaling Laws for Optimal Mixture-of-Experts Architecture Optimization
by: Wan, Weilin, et al.
Published: (2026)
by: Wan, Weilin, et al.
Published: (2026)
When Flat Minima Fail: Characterizing INT4 Quantization Collapse After FP32 Convergence
by: Armstrong, Marcus
Published: (2026)
by: Armstrong, Marcus
Published: (2026)
Step-Opt: Boosting Optimization Modeling in LLMs through Iterative Data Synthesis and Structured Validation
by: Wu, Yang, et al.
Published: (2025)
by: Wu, Yang, et al.
Published: (2025)
Noise Stability Optimization for Finding Flat Minima: A Hessian-based Regularization Approach
by: Zhang, Hongyang R., et al.
Published: (2023)
by: Zhang, Hongyang R., et al.
Published: (2023)
When Human Preferences Flip: An Instance-Dependent Robust Loss for RLHF
by: Xu, Yifan, et al.
Published: (2025)
by: Xu, Yifan, et al.
Published: (2025)
The Surprising Harmfulness of Benign Overfitting for Adversarial Robustness
by: Hao, Yifan, et al.
Published: (2024)
by: Hao, Yifan, et al.
Published: (2024)
Revisiting Energy-Based Model for Out-of-Distribution Detection
by: Wu, Yifan, et al.
Published: (2024)
by: Wu, Yifan, et al.
Published: (2024)
C-Flat++: Towards a More Efficient and Powerful Framework for Continual Learning
by: Li, Wei, et al.
Published: (2025)
by: Li, Wei, et al.
Published: (2025)
Identifying and Correcting Label Noise for Robust GNNs via Influence Contradiction
by: Ju, Wei, et al.
Published: (2026)
by: Ju, Wei, et al.
Published: (2026)
Explore and Establish Synergistic Effects Between Weight Pruning and Coreset Selection in Neural Network Training
by: Wan, Weilin, et al.
Published: (2025)
by: Wan, Weilin, et al.
Published: (2025)
Spurious Feature Diversification Improves Out-of-distribution Generalization
by: Lin, Yong, et al.
Published: (2023)
by: Lin, Yong, et al.
Published: (2023)
Towards Robust Content Watermarking Against Removal and Forgery Attacks
by: Zhu, Yifan, et al.
Published: (2026)
by: Zhu, Yifan, et al.
Published: (2026)
Local Minima Structures in Gaussian Mixture Models
by: Chen, Yudong, et al.
Published: (2020)
by: Chen, Yudong, et al.
Published: (2020)
DP-FedPGN: Finding Global Flat Minima for Differentially Private Federated Learning via Penalizing Gradient Norm
by: Liu, Junkang, et al.
Published: (2025)
by: Liu, Junkang, et al.
Published: (2025)
Beyond Sharp Minima: Robust LLM Unlearning via Feedback-Guided Multi-Point Optimization
by: Wu, Wenhan, et al.
Published: (2025)
by: Wu, Wenhan, et al.
Published: (2025)
Towards Efficient Automatic Self-Pruning of Large Language Models
by: Huang, Weizhong, et al.
Published: (2025)
by: Huang, Weizhong, et al.
Published: (2025)
New affine invariant ensemble samplers and their dimensional scaling
by: Chen, Yifan
Published: (2025)
by: Chen, Yifan
Published: (2025)
HGCN2SP: Hierarchical Graph Convolutional Network for Two-Stage Stochastic Programming
by: Wu, Yang, et al.
Published: (2025)
by: Wu, Yang, et al.
Published: (2025)
Bayesian Risk-Sensitive Policy Optimization For MDPs With General Loss Functions
by: Wang, Xiaoshuang, et al.
Published: (2025)
by: Wang, Xiaoshuang, et al.
Published: (2025)
MONAS: Efficient Zero-Shot Neural Architecture Search for MCUs
by: Qiao, Ye, et al.
Published: (2024)
by: Qiao, Ye, et al.
Published: (2024)
Nexus: Same Pretraining Loss, Better Downstream Generalization via Common Minima
by: Chen, Huanran, et al.
Published: (2026)
by: Chen, Huanran, et al.
Published: (2026)
Rethinking Graph Domain Adaptation: A Spectral Contrastive Perspective
by: Zhang, Haoyu, et al.
Published: (2025)
by: Zhang, Haoyu, et al.
Published: (2025)
An improved wind power prediction via a novel wind ramp identification algorithm
by: Xu, Yifan
Published: (2025)
by: Xu, Yifan
Published: (2025)
Instruction Tuning Changes How Upstream State Conditions Late Readout: A Cross-Patching Diagnostic
by: Zhou, Yifan
Published: (2026)
by: Zhou, Yifan
Published: (2026)
The Convergence Gap: Instruction-Tuned Language Models Stabilize Later in the Forward Pass
by: Zhou, Yifan
Published: (2026)
by: Zhou, Yifan
Published: (2026)
Similar Items
-
Optimized Gradient Clipping for Noisy Label Learning
by: Ye, Xichen, et al.
Published: (2024) -
Towards the Connection between Activation Sparsity and Flat Minima
by: Peng, Ze, et al.
Published: (2026) -
Are Flat Minima an Illusion?
by: Bennett, Michael Timothy
Published: (2026) -
A Flat Minima Perspective on Understanding Augmentations and Model Robustness
by: Yoo, Weebum, et al.
Published: (2025) -
Mirror Gradient: Towards Robust Multimodal Recommender Systems via Exploring Flat Local Minima
by: Zhong, Shanshan, et al.
Published: (2024)