Score-based Generative Models with Adaptive Momentum
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wen, Ziqing, Deng, Xiaoge, Luo, Ping, Sun, Tao, Li, Dongsheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Federated Prediction-Powered Inference from Decentralized Data
von: Luo, Ping, et al.
Veröffentlicht: (2024)
von: Luo, Ping, et al.
Veröffentlicht: (2024)
Accelerating Federated Learning by Selecting Beneficial Herd of Local Gradients
von: Luo, Ping, et al.
Veröffentlicht: (2024)
von: Luo, Ping, et al.
Veröffentlicht: (2024)
Stability and Generalization for Decentralized Markov SGD
von: Wang, Jiahuan, et al.
Veröffentlicht: (2026)
von: Wang, Jiahuan, et al.
Veröffentlicht: (2026)
Unveiling High-Probability Generalization in Decentralized SGD
von: Wang, Jiahuan, et al.
Veröffentlicht: (2026)
von: Wang, Jiahuan, et al.
Veröffentlicht: (2026)
FOAM: Blocked State Folding for Memory-Efficient LLM Training
von: Wen, Ziqing, et al.
Veröffentlicht: (2025)
von: Wen, Ziqing, et al.
Veröffentlicht: (2025)
GWT: Scalable Optimizer State Compression for Large Language Model Training
von: Wen, Ziqing, et al.
Veröffentlicht: (2025)
von: Wen, Ziqing, et al.
Veröffentlicht: (2025)
Local Gradient Regulation Stabilizes Federated Learning under Client Heterogeneity
von: Luo, Ping, et al.
Veröffentlicht: (2026)
von: Luo, Ping, et al.
Veröffentlicht: (2026)
Towards Understanding the Generalizability of Delayed Stochastic Gradient Descent
von: Deng, Xiaoge, et al.
Veröffentlicht: (2023)
von: Deng, Xiaoge, et al.
Veröffentlicht: (2023)
Sharpness-Aware Minimization with Adaptive Regularization for Training Deep Neural Networks
von: Zou, Jinping, et al.
Veröffentlicht: (2024)
von: Zou, Jinping, et al.
Veröffentlicht: (2024)
Revealing Modular Gradient Noise Imbalance in LLMs: Calibrating Adam via Signal-to-Noise Ratio
von: Wen, Ziqing, et al.
Veröffentlicht: (2026)
von: Wen, Ziqing, et al.
Veröffentlicht: (2026)
MLFEF: Machine Learning Fusion Model with Empirical Formula to Explore the Momentum in Competitive Sports
von: Peng, Ruixin, et al.
Veröffentlicht: (2024)
von: Peng, Ruixin, et al.
Veröffentlicht: (2024)
Optimizing the Adversarial Perturbation with a Momentum-based Adaptive Matrix
von: Tao, Wei, et al.
Veröffentlicht: (2025)
von: Tao, Wei, et al.
Veröffentlicht: (2025)
Communication-Efficient Distributed Learning via Sparse and Adaptive Stochastic Gradient
von: Deng, Xiaoge, et al.
Veröffentlicht: (2021)
von: Deng, Xiaoge, et al.
Veröffentlicht: (2021)
Variational Schrödinger Momentum Diffusion
von: Rojas, Kevin, et al.
Veröffentlicht: (2025)
von: Rojas, Kevin, et al.
Veröffentlicht: (2025)
Stability and Generalization for Stochastic Recursive Momentum-based Algorithms for (Strongly-)Convex One to $K$-Level Stochastic Optimizations
von: Pan, Xiaokang, et al.
Veröffentlicht: (2024)
von: Pan, Xiaokang, et al.
Veröffentlicht: (2024)
On the Performance Analysis of Momentum Method: A Frequency Domain Perspective
von: Li, Xianliang, et al.
Veröffentlicht: (2024)
von: Li, Xianliang, et al.
Veröffentlicht: (2024)
Approximation and Generalization Abilities of Score-based Neural Network Generative Models for Sub-Gaussian Distributions
von: Fu, Guoji, et al.
Veröffentlicht: (2025)
von: Fu, Guoji, et al.
Veröffentlicht: (2025)
Trivialized Momentum Facilitates Diffusion Generative Modeling on Lie Groups
von: Zhu, Yuchen, et al.
Veröffentlicht: (2024)
von: Zhu, Yuchen, et al.
Veröffentlicht: (2024)
Entropy-Based Dimension-Free Convergence and Loss-Adaptive Schedules for Diffusion Models
von: Aghapour, Ahmad, et al.
Veröffentlicht: (2026)
von: Aghapour, Ahmad, et al.
Veröffentlicht: (2026)
Incorporating Inductive Biases to Energy-based Generative Models
von: Li, Yukun, et al.
Veröffentlicht: (2025)
von: Li, Yukun, et al.
Veröffentlicht: (2025)
Adan: Adaptive Nesterov Momentum Algorithm for Faster Optimizing Deep Models
von: Xie, Xingyu, et al.
Veröffentlicht: (2022)
von: Xie, Xingyu, et al.
Veröffentlicht: (2022)
LLMExplainer: Large Language Model based Bayesian Inference for Graph Explanation Generation
von: Zhang, Jiaxing, et al.
Veröffentlicht: (2024)
von: Zhang, Jiaxing, et al.
Veröffentlicht: (2024)
TrasMuon: Trust-Region Adaptive Scaling for Orthogonalized Momentum Optimizers
von: Cheng, Peng, et al.
Veröffentlicht: (2026)
von: Cheng, Peng, et al.
Veröffentlicht: (2026)
Gaining Momentum: Uncovering Hidden Scoring Dynamics in Hockey through Deep Neural Sequencing and Causal Modeling
von: Griffiths, Daniel, et al.
Veröffentlicht: (2025)
von: Griffiths, Daniel, et al.
Veröffentlicht: (2025)
Score-based Generative Modeling for Conditional Independence Testing
von: Ren, Yixin, et al.
Veröffentlicht: (2025)
von: Ren, Yixin, et al.
Veröffentlicht: (2025)
Joint Selection for Large-Scale Pre-Training Data via Policy Gradient-based Mask Learning
von: Fan, Ziqing, et al.
Veröffentlicht: (2025)
von: Fan, Ziqing, et al.
Veröffentlicht: (2025)
Variance-Adaptive Muon: Accelerating LLM Pretraining with NSR-Modulated and Variance-Scaled Momentum
von: Li, Jingru, et al.
Veröffentlicht: (2026)
von: Li, Jingru, et al.
Veröffentlicht: (2026)
Where to Spend Rollouts: Hit-Utility Optimal Rollout Allocation for Group-Based RLVR
von: Wang, Tao, et al.
Veröffentlicht: (2026)
von: Wang, Tao, et al.
Veröffentlicht: (2026)
Continual Task Learning through Adaptive Policy Self-Composition
von: Hu, Shengchao, et al.
Veröffentlicht: (2024)
von: Hu, Shengchao, et al.
Veröffentlicht: (2024)
Moderating the Generalization of Score-based Generative Model
von: Jiang, Wan, et al.
Veröffentlicht: (2024)
von: Jiang, Wan, et al.
Veröffentlicht: (2024)
Adaptive Memory Momentum via a Model-Based Framework for Deep Learning Optimization
von: Topollai, Kristi, et al.
Veröffentlicht: (2025)
von: Topollai, Kristi, et al.
Veröffentlicht: (2025)
Generalizable and Efficient Automated Scoring with a Knowledge-Distilled Multi-Task Mixture-of-Experts
von: Fang, Luyang, et al.
Veröffentlicht: (2025)
von: Fang, Luyang, et al.
Veröffentlicht: (2025)
Provable Generalization Bounds for Deep Neural Networks with Momentum-Adaptive Gradient Dropout
von: Safder, Adeel
Veröffentlicht: (2025)
von: Safder, Adeel
Veröffentlicht: (2025)
TSGM: Regular and Irregular Time-series Generation using Score-based Generative Models
von: Lim, Haksoo, et al.
Veröffentlicht: (2025)
von: Lim, Haksoo, et al.
Veröffentlicht: (2025)
ScoreFusion: Fusing Score-based Generative Models via Kullback-Leibler Barycenters
von: Liu, Hao, et al.
Veröffentlicht: (2024)
von: Liu, Hao, et al.
Veröffentlicht: (2024)
Intrinsic Wasserstein Rates for Score-Based Generative Models on Smooth Manifolds
von: Fu, Guoji, et al.
Veröffentlicht: (2026)
von: Fu, Guoji, et al.
Veröffentlicht: (2026)
Multi-source Unsupervised Domain Adaptation on Graphs with Transferability Modeling
von: Zhao, Tianxiang, et al.
Veröffentlicht: (2024)
von: Zhao, Tianxiang, et al.
Veröffentlicht: (2024)
MambaJSCC: Adaptive Deep Joint Source-Channel Coding with Generalized State Space Model
von: Wu, Tong, et al.
Veröffentlicht: (2024)
von: Wu, Tong, et al.
Veröffentlicht: (2024)
On the Generalization of Stochastic Gradient Descent with Momentum
von: Ramezani-Kebrya, Ali, et al.
Veröffentlicht: (2018)
von: Ramezani-Kebrya, Ali, et al.
Veröffentlicht: (2018)
Feature-Function Curvature Analysis: A Geometric Framework for Explaining Differentiable Models
von: Najafi, Hamed, et al.
Veröffentlicht: (2025)
von: Najafi, Hamed, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Federated Prediction-Powered Inference from Decentralized Data
von: Luo, Ping, et al.
Veröffentlicht: (2024) -
Accelerating Federated Learning by Selecting Beneficial Herd of Local Gradients
von: Luo, Ping, et al.
Veröffentlicht: (2024) -
Stability and Generalization for Decentralized Markov SGD
von: Wang, Jiahuan, et al.
Veröffentlicht: (2026) -
Unveiling High-Probability Generalization in Decentralized SGD
von: Wang, Jiahuan, et al.
Veröffentlicht: (2026) -
FOAM: Blocked State Folding for Memory-Efficient LLM Training
von: Wen, Ziqing, et al.
Veröffentlicht: (2025)