Learning Gradient-based Mixup with Extrapolation toward Flatter Minima for Domain Generalization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Peng, Danni, Pan, Sinno Jialin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Gradient Descent with Polyak's Momentum Finds Flatter Minima via Large Catapults
von: Phunyaphibarn, Prin, et al.
Veröffentlicht: (2023)
von: Phunyaphibarn, Prin, et al.
Veröffentlicht: (2023)
Gradient Descent Converges Linearly to Flatter Minima than Gradient Flow in Shallow Linear Networks
von: Beneventano, Pierfrancesco, et al.
Veröffentlicht: (2025)
von: Beneventano, Pierfrancesco, et al.
Veröffentlicht: (2025)
Bilateral Sharpness-Aware Minimization for Flatter Minima
von: Deng, Jiaxin, et al.
Veröffentlicht: (2024)
von: Deng, Jiaxin, et al.
Veröffentlicht: (2024)
Sharpness-Aware Minimization Efficiently Selects Flatter Minima Late in Training
von: Zhou, Zhanpeng, et al.
Veröffentlicht: (2024)
von: Zhou, Zhanpeng, et al.
Veröffentlicht: (2024)
MetaDefense: Defending Finetuning-based Jailbreak Attack Before and During Generation
von: Jiang, Weisen, et al.
Veröffentlicht: (2025)
von: Jiang, Weisen, et al.
Veröffentlicht: (2025)
Fast Graph Generation via Spectral Diffusion
von: Luo, Tianze, et al.
Veröffentlicht: (2022)
von: Luo, Tianze, et al.
Veröffentlicht: (2022)
Improving the Generalization of Unseen Crowd Behaviors for Reinforcement Learning based Local Motion Planners
von: Ng, Wen Zheng Terence, et al.
Veröffentlicht: (2024)
von: Ng, Wen Zheng Terence, et al.
Veröffentlicht: (2024)
State Chrono Representation for Enhancing Generalization in Reinforcement Learning
von: Chen, Jianda, et al.
Veröffentlicht: (2024)
von: Chen, Jianda, et al.
Veröffentlicht: (2024)
Variational Learning Finds Flatter Solutions at the Edge of Stability
von: Ghosh, Avrajit, et al.
Veröffentlicht: (2025)
von: Ghosh, Avrajit, et al.
Veröffentlicht: (2025)
MetaMoE: Diversity-Aware Proxy Selection for Privacy-Preserving Mixture-of-Experts Unification
von: Jiang, Weisen, et al.
Veröffentlicht: (2026)
von: Jiang, Weisen, et al.
Veröffentlicht: (2026)
Overcoming Negative Transfer by Online Selection: Distant Domain Adaptation for Fault Diagnosis
von: Wang, Ziyan, et al.
Veröffentlicht: (2024)
von: Wang, Ziyan, et al.
Veröffentlicht: (2024)
Single Domain Generalization with Model-aware Parametric Batch-wise Mixup
von: Heidari, Marzi, et al.
Veröffentlicht: (2025)
von: Heidari, Marzi, et al.
Veröffentlicht: (2025)
Learning Fair Robustness via Domain Mixup
von: Zhong, Meiyu, et al.
Veröffentlicht: (2024)
von: Zhong, Meiyu, et al.
Veröffentlicht: (2024)
Adversarial Mixup Unlearning
von: Peng, Zhuoyi, et al.
Veröffentlicht: (2025)
von: Peng, Zhuoyi, et al.
Veröffentlicht: (2025)
Towards Optimization and Model Selection for Domain Generalization: A Mixup-guided Solution
von: Lu, Wang, et al.
Veröffentlicht: (2022)
von: Lu, Wang, et al.
Veröffentlicht: (2022)
Linearly Convergent Mixup Learning
von: Obi, Gakuto, et al.
Veröffentlicht: (2025)
von: Obi, Gakuto, et al.
Veröffentlicht: (2025)
Gradient Extrapolation for Debiased Representation Learning
von: Asaad, Ihab, et al.
Veröffentlicht: (2025)
von: Asaad, Ihab, et al.
Veröffentlicht: (2025)
Beyond Speedup -- Utilizing KV Cache for Sampling and Reasoning
von: Xing, Zeyu, et al.
Veröffentlicht: (2026)
von: Xing, Zeyu, et al.
Veröffentlicht: (2026)
On Mixup Regularization
von: Carratino, Luigi, et al.
Veröffentlicht: (2020)
von: Carratino, Luigi, et al.
Veröffentlicht: (2020)
GradMix: Gradient-based Selective Mixup for Robust Data Augmentation in Class-Incremental Learning
von: Kim, Minsu, et al.
Veröffentlicht: (2025)
von: Kim, Minsu, et al.
Veröffentlicht: (2025)
Gradient Extrapolation-Based Policy Optimization
von: Swapnil, Ismam Nur, et al.
Veröffentlicht: (2026)
von: Swapnil, Ismam Nur, et al.
Veröffentlicht: (2026)
Mixup Domain Adaptations for Dynamic Remaining Useful Life Predictions
von: Furqon, Muhammad Tanzil, et al.
Veröffentlicht: (2024)
von: Furqon, Muhammad Tanzil, et al.
Veröffentlicht: (2024)
RLPR: Extrapolating RLVR to General Domains without Verifiers
von: Yu, Tianyu, et al.
Veröffentlicht: (2025)
von: Yu, Tianyu, et al.
Veröffentlicht: (2025)
AuG-KD: Anchor-Based Mixup Generation for Out-of-Domain Knowledge Distillation
von: Tang, Zihao, et al.
Veröffentlicht: (2024)
von: Tang, Zihao, et al.
Veröffentlicht: (2024)
From Pruning to Grafting: Dynamic Knowledge Redistribution via Learnable Layer Fusion
von: Pei, Zehua, et al.
Veröffentlicht: (2024)
von: Pei, Zehua, et al.
Veröffentlicht: (2024)
Flat Minima and Generalization: Insights from Stochastic Convex Optimization
von: Schliserman, Matan, et al.
Veröffentlicht: (2025)
von: Schliserman, Matan, et al.
Veröffentlicht: (2025)
Learning to Extrapolate to New Tasks: A Relational Approach to Task Extrapolation
von: Ousherovitch, Adam, et al.
Veröffentlicht: (2026)
von: Ousherovitch, Adam, et al.
Veröffentlicht: (2026)
Robust Gradient Descent via Heavy-Ball Momentum with Predictive Extrapolation
von: Ali, Sarwan
Veröffentlicht: (2025)
von: Ali, Sarwan
Veröffentlicht: (2025)
Gradient-Guided Annealing for Domain Generalization
von: Ballas, Aristotelis, et al.
Veröffentlicht: (2025)
von: Ballas, Aristotelis, et al.
Veröffentlicht: (2025)
Augment on Manifold: Mixup Regularization with UMAP
von: El-Laham, Yousef, et al.
Veröffentlicht: (2023)
von: El-Laham, Yousef, et al.
Veröffentlicht: (2023)
Mixup Regularization: A Probabilistic Perspective
von: El-Laham, Yousef, et al.
Veröffentlicht: (2025)
von: El-Laham, Yousef, et al.
Veröffentlicht: (2025)
Are Flat Minima an Illusion?
von: Bennett, Michael Timothy
Veröffentlicht: (2026)
von: Bennett, Michael Timothy
Veröffentlicht: (2026)
Sharp Minima Can Generalize: A Loss Landscape Perspective On Data
von: Fan, Raymond, et al.
Veröffentlicht: (2025)
von: Fan, Raymond, et al.
Veröffentlicht: (2025)
DP-FedPGN: Finding Global Flat Minima for Differentially Private Federated Learning via Penalizing Gradient Norm
von: Liu, Junkang, et al.
Veröffentlicht: (2025)
von: Liu, Junkang, et al.
Veröffentlicht: (2025)
Towards the Connection between Activation Sparsity and Flat Minima
von: Peng, Ze, et al.
Veröffentlicht: (2026)
von: Peng, Ze, et al.
Veröffentlicht: (2026)
PreMoE: Proactive Inference for Efficient Mixture-of-Experts
von: Pei, Zehua, et al.
Veröffentlicht: (2025)
von: Pei, Zehua, et al.
Veröffentlicht: (2025)
SEDGE: Structural Extrapolated Data Generation
von: Zhang, Kun, et al.
Veröffentlicht: (2026)
von: Zhang, Kun, et al.
Veröffentlicht: (2026)
RAMPAGE: RAndomized Mid-Point for debiAsed Gradient Extrapolation
von: Luo, Zhankun, et al.
Veröffentlicht: (2026)
von: Luo, Zhankun, et al.
Veröffentlicht: (2026)
Seeing Beyond: Extrapolative Domain Adaptive Panoramic Segmentation
von: Zheng, Yuanfan, et al.
Veröffentlicht: (2026)
von: Zheng, Yuanfan, et al.
Veröffentlicht: (2026)
Mirror Gradient: Towards Robust Multimodal Recommender Systems via Exploring Flat Local Minima
von: Zhong, Shanshan, et al.
Veröffentlicht: (2024)
von: Zhong, Shanshan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Gradient Descent with Polyak's Momentum Finds Flatter Minima via Large Catapults
von: Phunyaphibarn, Prin, et al.
Veröffentlicht: (2023) -
Gradient Descent Converges Linearly to Flatter Minima than Gradient Flow in Shallow Linear Networks
von: Beneventano, Pierfrancesco, et al.
Veröffentlicht: (2025) -
Bilateral Sharpness-Aware Minimization for Flatter Minima
von: Deng, Jiaxin, et al.
Veröffentlicht: (2024) -
Sharpness-Aware Minimization Efficiently Selects Flatter Minima Late in Training
von: Zhou, Zhanpeng, et al.
Veröffentlicht: (2024) -
MetaDefense: Defending Finetuning-based Jailbreak Attack Before and During Generation
von: Jiang, Weisen, et al.
Veröffentlicht: (2025)