Offline Policy Learning with Weight Clipping and Heaviside Composite Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Jingren, Qin, Hanzhang, Liu, Junyi, Chou, Mabel C., Pang, Jong-Shi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Online Resource Allocation with Non-Stationary Customers
von: Zhang, Xiaoyue, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoyue, et al.
Veröffentlicht: (2024)
Classification and Treatment Learning with Constraints via Composite Heaviside Optimization: a Progressive MIP Method
von: Fang, Yue, et al.
Veröffentlicht: (2024)
von: Fang, Yue, et al.
Veröffentlicht: (2024)
Solving Constrained Affine Heaviside Composite Optimization Problems by a Progressive IP Approach
von: Zheng, Ke, et al.
Veröffentlicht: (2026)
von: Zheng, Ke, et al.
Veröffentlicht: (2026)
Heaviside Low-Rank Support Matrix Machine
von: Xiu, Xianchao, et al.
Veröffentlicht: (2026)
von: Xiu, Xianchao, et al.
Veröffentlicht: (2026)
Clipped Gradient Methods for Nonsmooth Convex Optimization under Heavy-Tailed Noise: A Refined Analysis
von: Liu, Zijian
Veröffentlicht: (2025)
von: Liu, Zijian
Veröffentlicht: (2025)
Nonconvex Stochastic Optimization under Heavy-Tailed Noises: Optimal Convergence without Gradient Clipping
von: Liu, Zijian, et al.
Veröffentlicht: (2024)
von: Liu, Zijian, et al.
Veröffentlicht: (2024)
Non-Asymptotic Global Convergence of PPO-Clip
von: Liu, Yin, et al.
Veröffentlicht: (2025)
von: Liu, Yin, et al.
Veröffentlicht: (2025)
Adaptive Learning-based Surrogate Method for Stochastic Programs with Implicitly Decision-dependent Uncertainty
von: Shen, Boyang, et al.
Veröffentlicht: (2025)
von: Shen, Boyang, et al.
Veröffentlicht: (2025)
Offline Reinforcement Learning via Inverse Optimization
von: Dimanidis, Ioannis, et al.
Veröffentlicht: (2025)
von: Dimanidis, Ioannis, et al.
Veröffentlicht: (2025)
Compressed Proximal Federated Learning for Non-Convex Composite Optimization on Heterogeneous Data
von: Qiu, Pu, et al.
Veröffentlicht: (2026)
von: Qiu, Pu, et al.
Veröffentlicht: (2026)
Offline Hierarchical Reinforcement Learning via Inverse Optimization
von: Schmidt, Carolin, et al.
Veröffentlicht: (2024)
von: Schmidt, Carolin, et al.
Veröffentlicht: (2024)
Differentially Private Clipped-SGD: High-Probability Convergence with Arbitrary Clipping Level
von: Khah, Saleh Vatan, et al.
Veröffentlicht: (2025)
von: Khah, Saleh Vatan, et al.
Veröffentlicht: (2025)
Tight Long-Term Tail Decay of (Clipped) SGD in Non-Convex Optimization
von: Armacki, Aleksandar, et al.
Veröffentlicht: (2026)
von: Armacki, Aleksandar, et al.
Veröffentlicht: (2026)
Methods for Convex $(L_0,L_1)$-Smooth Optimization: Clipping, Acceleration, and Adaptivity
von: Gorbunov, Eduard, et al.
Veröffentlicht: (2024)
von: Gorbunov, Eduard, et al.
Veröffentlicht: (2024)
Robust Implicit Regularization via Weight Normalization
von: Chou, Hung-Hsu, et al.
Veröffentlicht: (2023)
von: Chou, Hung-Hsu, et al.
Veröffentlicht: (2023)
Statistical Robustness of Interval CVaR Based Regression Models under Perturbation and Contamination
von: You, Yulei, et al.
Veröffentlicht: (2026)
von: You, Yulei, et al.
Veröffentlicht: (2026)
Revisiting Gradient Normalization and Clipping for Nonconvex SGD under Heavy-Tailed Noise: Necessity, Sufficiency, and Acceleration
von: Sun, Tao, et al.
Veröffentlicht: (2024)
von: Sun, Tao, et al.
Veröffentlicht: (2024)
Second-order Optimization under Heavy-Tailed Noise: Hessian Clipping and Sample Complexity Limits
von: Sadiev, Abdurakhmon, et al.
Veröffentlicht: (2025)
von: Sadiev, Abdurakhmon, et al.
Veröffentlicht: (2025)
Provably Faster Algorithms for Bilevel Optimization via Without-Replacement Sampling
von: Li, Junyi, et al.
Veröffentlicht: (2024)
von: Li, Junyi, et al.
Veröffentlicht: (2024)
Double Duality: Variational Primal-Dual Policy Optimization for Constrained Reinforcement Learning
von: Li, Zihao, et al.
Veröffentlicht: (2024)
von: Li, Zihao, et al.
Veröffentlicht: (2024)
Convergence of Clipped-SGD for Convex $(L_0,L_1)$-Smooth Optimization with Heavy-Tailed Noise
von: Chezhegov, Savelii, et al.
Veröffentlicht: (2025)
von: Chezhegov, Savelii, et al.
Veröffentlicht: (2025)
A Weighted Gradient Tracking Privacy-Preserving Method for Distributed Optimization
von: Xie, Furan, et al.
Veröffentlicht: (2025)
von: Xie, Furan, et al.
Veröffentlicht: (2025)
Residuals-based Offline Reinforcement Learning
von: Zhu, Qing, et al.
Veröffentlicht: (2026)
von: Zhu, Qing, et al.
Veröffentlicht: (2026)
Robust and Fast Training via Per-Sample Clipping
von: Nobile, Davide, et al.
Veröffentlicht: (2026)
von: Nobile, Davide, et al.
Veröffentlicht: (2026)
From Gradient Clipping to Normalization for Heavy Tailed SGD
von: Hübler, Florian, et al.
Veröffentlicht: (2024)
von: Hübler, Florian, et al.
Veröffentlicht: (2024)
Parameter-free Clipped Gradient Descent Meets Polyak
von: Takezawa, Yuki, et al.
Veröffentlicht: (2024)
von: Takezawa, Yuki, et al.
Veröffentlicht: (2024)
Online Residual Learning from Offline Experts for Pedestrian Tracking
von: Vlachos, Anastasios, et al.
Veröffentlicht: (2024)
von: Vlachos, Anastasios, et al.
Veröffentlicht: (2024)
Data-driven Piecewise Affine Decision Rules for Stochastic Programming with Covariate Information
von: Zhang, Yiyang, et al.
Veröffentlicht: (2023)
von: Zhang, Yiyang, et al.
Veröffentlicht: (2023)
Clipped SGD Algorithms for Performative Prediction: Tight Bounds for Clipping Bias and Remedies
von: Li, Qiang, et al.
Veröffentlicht: (2024)
von: Li, Qiang, et al.
Veröffentlicht: (2024)
Regularized Gradient Clipping Provably Trains Wide and Deep Neural Networks
von: Tucat, Matteo, et al.
Veröffentlicht: (2024)
von: Tucat, Matteo, et al.
Veröffentlicht: (2024)
Double Momentum and Error Feedback for Clipping with Fast Rates and Differential Privacy
von: Islamov, Rustem, et al.
Veröffentlicht: (2025)
von: Islamov, Rustem, et al.
Veröffentlicht: (2025)
Efficient Private SCO for Heavy-Tailed Data via Averaged Clipping
von: Jin, Chenhan, et al.
Veröffentlicht: (2022)
von: Jin, Chenhan, et al.
Veröffentlicht: (2022)
A Compositional Kernel Model for Feature Learning
von: Ruan, Feng, et al.
Veröffentlicht: (2025)
von: Ruan, Feng, et al.
Veröffentlicht: (2025)
Adaptive Federated Minimax Optimization with Lower Complexities
von: Huang, Feihu, et al.
Veröffentlicht: (2022)
von: Huang, Feihu, et al.
Veröffentlicht: (2022)
Offline-Online Reinforcement Learning for Linear Mixture MDPs
von: Zhang, Zhongjun, et al.
Veröffentlicht: (2026)
von: Zhang, Zhongjun, et al.
Veröffentlicht: (2026)
Reward-Relevance-Filtered Linear Offline Reinforcement Learning
von: Zhou, Angela
Veröffentlicht: (2024)
von: Zhou, Angela
Veröffentlicht: (2024)
Stochastic Compositional Optimization with Compositional Constraints
von: Yang, Shuoguang, et al.
Veröffentlicht: (2022)
von: Yang, Shuoguang, et al.
Veröffentlicht: (2022)
Operator Models for Continuous-Time Offline Reinforcement Learning
von: Hoischen, Nicolas, et al.
Veröffentlicht: (2025)
von: Hoischen, Nicolas, et al.
Veröffentlicht: (2025)
Achieving Linear Speedup for Composite Federated Learning
von: Huang, Kun, et al.
Veröffentlicht: (2026)
von: Huang, Kun, et al.
Veröffentlicht: (2026)
Gradient Clipping Beyond Vector Norms: A Spectral Approach for Matrix-Valued Parameters
von: Yukhimchuk, Alexander, et al.
Veröffentlicht: (2026)
von: Yukhimchuk, Alexander, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Online Resource Allocation with Non-Stationary Customers
von: Zhang, Xiaoyue, et al.
Veröffentlicht: (2024) -
Classification and Treatment Learning with Constraints via Composite Heaviside Optimization: a Progressive MIP Method
von: Fang, Yue, et al.
Veröffentlicht: (2024) -
Solving Constrained Affine Heaviside Composite Optimization Problems by a Progressive IP Approach
von: Zheng, Ke, et al.
Veröffentlicht: (2026) -
Heaviside Low-Rank Support Matrix Machine
von: Xiu, Xianchao, et al.
Veröffentlicht: (2026) -
Clipped Gradient Methods for Nonsmooth Convex Optimization under Heavy-Tailed Noise: A Refined Analysis
von: Liu, Zijian
Veröffentlicht: (2025)