Saved in:
| Main Authors: | Sun, Zhongchang, He, Sihong, Miao, Fei, Zou, Shaofeng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2405.01327 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Momentum for the Win: Collaborative Federated Reinforcement Learning across Heterogeneous Environments
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
CUQDS: Conformal Uncertainty Quantification under Distribution Shift for Trajectory Prediction
by: Huang, Huiqun, et al.
Published: (2024)
by: Huang, Huiqun, et al.
Published: (2024)
Model-Free Robust Reinforcement Learning with Sample Complexity Analysis
by: Wang, Yudan, et al.
Published: (2024)
by: Wang, Yudan, et al.
Published: (2024)
Near-Optimal Sample Complexity for Iterated CVaR Reinforcement Learning with a Generative Model
by: Deng, Zilong, et al.
Published: (2025)
by: Deng, Zilong, et al.
Published: (2025)
HIPO: Instruction Hierarchy via Constrained Reinforcement Learning
by: Chen, Keru, et al.
Published: (2026)
by: Chen, Keru, et al.
Published: (2026)
Achieving the Asymptotically Optimal Sample Complexity of Offline Reinforcement Learning: A DRO-Based Approach
by: Wang, Yue, et al.
Published: (2023)
by: Wang, Yue, et al.
Published: (2023)
Low-Rank Adaptation for Critic Learning in Off-Policy Reinforcement Learning
by: Zhuang, Yuan, et al.
Published: (2026)
by: Zhuang, Yuan, et al.
Published: (2026)
Robust Multi-Hypothesis Testing with Moment Constrained Uncertainty Sets
by: Magesh, Akshayaa, et al.
Published: (2022)
by: Magesh, Akshayaa, et al.
Published: (2022)
Theoretical Study of Conflict-Avoidant Multi-Objective Reinforcement Learning
by: Wang, Yudan, et al.
Published: (2024)
by: Wang, Yudan, et al.
Published: (2024)
Detector-Evasive LLM Paraphrasing via Constrained Policy Optimization
by: Wang, Mingyi, et al.
Published: (2026)
by: Wang, Mingyi, et al.
Published: (2026)
Understanding Uncertainty-based Active Learning Under Model Mismatch
by: Rahmati, Amir Hossein, et al.
Published: (2024)
by: Rahmati, Amir Hossein, et al.
Published: (2024)
Neural Stochastic Differential Equations with Change Points: A Generative Adversarial Approach
by: Sun, Zhongchang, et al.
Published: (2023)
by: Sun, Zhongchang, et al.
Published: (2023)
What is the Solution for State-Adversarial Multi-Agent Reinforcement Learning?
by: Han, Songyang, et al.
Published: (2022)
by: Han, Songyang, et al.
Published: (2022)
Pessimism Principle Can Be Effective: Towards a Framework for Zero-Shot Transfer Reinforcement Learning
by: Zhang, Chi, et al.
Published: (2025)
by: Zhang, Chi, et al.
Published: (2025)
Large-Scale Non-convex Stochastic Constrained Distributionally Robust Optimization
by: Zhang, Qi, et al.
Published: (2024)
by: Zhang, Qi, et al.
Published: (2024)
RLVR-World: Training World Models with Reinforcement Learning
by: Wu, Jialong, et al.
Published: (2025)
by: Wu, Jialong, et al.
Published: (2025)
A Unified View on Solving Objective Mismatch in Model-Based Reinforcement Learning
by: Wei, Ran, et al.
Published: (2023)
by: Wei, Ran, et al.
Published: (2023)
Quantile Geometry Regularization for Distributional Reinforcement Learning
by: Zhang, Zhaofan, et al.
Published: (2026)
by: Zhang, Zhaofan, et al.
Published: (2026)
Multi-Agent Deep Reinforcement Learning Under Constrained Communications
by: Shaik, Shahil, et al.
Published: (2026)
by: Shaik, Shahil, et al.
Published: (2026)
LDC-MTL: Balancing Multi-Task Learning through Scalable Loss Discrepancy Control
by: Xiao, Peiyao, et al.
Published: (2025)
by: Xiao, Peiyao, et al.
Published: (2025)
Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning
by: Chen, Rufeng, et al.
Published: (2026)
by: Chen, Rufeng, et al.
Published: (2026)
Finite-Time Error Bounds for Greedy-GQ
by: Wang, Yue, et al.
Published: (2022)
by: Wang, Yue, et al.
Published: (2022)
The Alignment Ceiling: Objective Mismatch in Reinforcement Learning from Human Feedback
by: Lambert, Nathan, et al.
Published: (2023)
by: Lambert, Nathan, et al.
Published: (2023)
Variational Neural Stochastic Differential Equations with Change Points
by: El-Laham, Yousef, et al.
Published: (2024)
by: El-Laham, Yousef, et al.
Published: (2024)
Diagnosing Training Inference Mismatch in LLM Reinforcement Learning
by: Zhong, Tianle, et al.
Published: (2026)
by: Zhong, Tianle, et al.
Published: (2026)
Are Large Language Models Chameleons? An Attempt to Simulate Social Surveys
by: Geng, Mingmeng, et al.
Published: (2024)
by: Geng, Mingmeng, et al.
Published: (2024)
Convergence Guarantees for RMSProp and Adam in Generalized-smooth Non-convex Optimization with Affine Noise Variance
by: Zhang, Qi, et al.
Published: (2024)
by: Zhang, Qi, et al.
Published: (2024)
Why Adam Can Beat SGD: Second-Moment Normalization Yields Sharper Tails
by: Jin, Ruinan, et al.
Published: (2026)
by: Jin, Ruinan, et al.
Published: (2026)
DiffCPS: Diffusion Model based Constrained Policy Search for Offline Reinforcement Learning
by: He, Longxiang, et al.
Published: (2023)
by: He, Longxiang, et al.
Published: (2023)
Robust Conformal Prediction under Distribution Shift via Physics-Informed Structural Causal Model
by: Xu, Rui, et al.
Published: (2024)
by: Xu, Rui, et al.
Published: (2024)
Collaborative Value Function Estimation Under Model Mismatch: A Federated Temporal Difference Analysis
by: Beikmohammadi, Ali, et al.
Published: (2025)
by: Beikmohammadi, Ali, et al.
Published: (2025)
Sample Complexity Characterization for Linear Contextual MDPs
by: Deng, Junze, et al.
Published: (2024)
by: Deng, Junze, et al.
Published: (2024)
Adaptive Gradient Normalization and Independent Sampling for (Stochastic) Generalized-Smooth Optimization
by: Yang, Yufeng, et al.
Published: (2024)
by: Yang, Yufeng, et al.
Published: (2024)
Sparsity-based Safety Conservatism for Constrained Offline Reinforcement Learning
by: Cho, Minjae, et al.
Published: (2024)
by: Cho, Minjae, et al.
Published: (2024)
Secure Resource Allocation via Constrained Deep Reinforcement Learning
by: Sun, Jianfei, et al.
Published: (2025)
by: Sun, Jianfei, et al.
Published: (2025)
Attribution-Guided Continual Learning for Large Language Models
by: Liu, Yazheng, et al.
Published: (2026)
by: Liu, Yazheng, et al.
Published: (2026)
Constrained Meta Agnostic Reinforcement Learning
by: Daaboul, Karam, et al.
Published: (2024)
by: Daaboul, Karam, et al.
Published: (2024)
An End-to-End Reinforcement Learning Based Approach for Micro-View Order-Dispatching in Ride-Hailing
by: Yue, Xinlang, et al.
Published: (2024)
by: Yue, Xinlang, et al.
Published: (2024)
Planning Under Observation Mismatch for Traffic Signal Control via Adaptive Modular World Models
by: Huang, Zherui, et al.
Published: (2025)
by: Huang, Zherui, et al.
Published: (2025)
Improving RCT-Based CATE Estimation Under Covariate Mismatch via Calibrated Alignment
by: Asiaee, Amir, et al.
Published: (2026)
by: Asiaee, Amir, et al.
Published: (2026)
Similar Items
-
Momentum for the Win: Collaborative Federated Reinforcement Learning across Heterogeneous Environments
by: Wang, Han, et al.
Published: (2024) -
CUQDS: Conformal Uncertainty Quantification under Distribution Shift for Trajectory Prediction
by: Huang, Huiqun, et al.
Published: (2024) -
Model-Free Robust Reinforcement Learning with Sample Complexity Analysis
by: Wang, Yudan, et al.
Published: (2024) -
Near-Optimal Sample Complexity for Iterated CVaR Reinforcement Learning with a Generative Model
by: Deng, Zilong, et al.
Published: (2025) -
HIPO: Instruction Hierarchy via Constrained Reinforcement Learning
by: Chen, Keru, et al.
Published: (2026)