State-wise Constrained Policy Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Weiye, Chen, Rui, Sun, Yifan, Wei, Tianhao, Liu, Changliu |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learn With Imagination: Safe Set Guided State-wise Constrained Policy Optimization
by: Sun, Yifan, et al.
Published: (2023)
by: Sun, Yifan, et al.
Published: (2023)
Absolute Policy Optimization
by: Zhao, Weiye, et al.
Published: (2023)
by: Zhao, Weiye, et al.
Published: (2023)
Absolute State-wise Constrained Policy Optimization: High-Probability State-wise Constraints Satisfaction
by: Zhao, Weiye, et al.
Published: (2024)
by: Zhao, Weiye, et al.
Published: (2024)
GUARD: A Safe Reinforcement Learning Benchmark
by: Zhao, Weiye, et al.
Published: (2023)
by: Zhao, Weiye, et al.
Published: (2023)
Physics-Aware Combinatorial Assembly Sequence Planning using Data-free Action Masking
by: Liu, Ruixuan, et al.
Published: (2024)
by: Liu, Ruixuan, et al.
Published: (2024)
Implicit Safe Set Algorithm for Provably Safe Reinforcement Learning
by: Zhao, Weiye, et al.
Published: (2024)
by: Zhao, Weiye, et al.
Published: (2024)
A Lightweight and Transferable Design for Robust LEGO Manipulation
by: Liu, Ruixuan, et al.
Published: (2023)
by: Liu, Ruixuan, et al.
Published: (2023)
Meta-Control: Automatic Model-based Control Synthesis for Heterogeneous Robot Skills
by: Wei, Tianhao, et al.
Published: (2024)
by: Wei, Tianhao, et al.
Published: (2024)
Verification of Neural Control Barrier Functions with Symbolic Derivative Bounds Propagation
by: Hu, Hanjiang, et al.
Published: (2024)
by: Hu, Hanjiang, et al.
Published: (2024)
Continual Learning and Lifting of Koopman Dynamics for Linear Control of Legged Robots
by: Li, Feihan, et al.
Published: (2024)
by: Li, Feihan, et al.
Published: (2024)
Estimating Neural Network Robustness via Lipschitz Constant and Architecture Sensitivity
by: Abuduweili, Abulikemu, et al.
Published: (2024)
by: Abuduweili, Abulikemu, et al.
Published: (2024)
Constrained Decoding for Safe Robot Navigation Foundation Models
by: Kapoor, Parv, et al.
Published: (2025)
by: Kapoor, Parv, et al.
Published: (2025)
Constrained Group Relative Policy Optimization
by: Girgis, Roger, et al.
Published: (2026)
by: Girgis, Roger, et al.
Published: (2026)
Constrained Policy Optimization via Sampling-Based Weight-Space Projection
by: Cao, Shengfan, et al.
Published: (2025)
by: Cao, Shengfan, et al.
Published: (2025)
Diffusion Policy through Conditional Proximal Policy Optimization
by: Liu, Ben, et al.
Published: (2026)
by: Liu, Ben, et al.
Published: (2026)
Dexterous Safe Control for Humanoids in Cluttered Environments via Projected Safe Set Algorithm
by: Chen, Rui, et al.
Published: (2025)
by: Chen, Rui, et al.
Published: (2025)
PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning
by: Yang, Shunpeng, et al.
Published: (2026)
by: Yang, Shunpeng, et al.
Published: (2026)
Evolutionary Policy Optimization
by: Wang, Jianren, et al.
Published: (2025)
by: Wang, Jianren, et al.
Published: (2025)
From Decoupled to Coupled: Robustness Verification for Learning-based Keypoint Detection with Joint Specifications
by: Luo, Xusheng, et al.
Published: (2026)
by: Luo, Xusheng, et al.
Published: (2026)
Real-Time Safe Control of Neural Network Dynamic Models with Sound Approximation
by: Hu, Hanjiang, et al.
Published: (2024)
by: Hu, Hanjiang, et al.
Published: (2024)
Safe PDE Boundary Control with Neural Operators
by: Hu, Hanjiang, et al.
Published: (2024)
by: Hu, Hanjiang, et al.
Published: (2024)
Dichotomous Diffusion Policy Optimization
by: Liang, Ruiming, et al.
Published: (2025)
by: Liang, Ruiming, et al.
Published: (2025)
Diffusion Policy Policy Optimization
by: Ren, Allen Z., et al.
Published: (2024)
by: Ren, Allen Z., et al.
Published: (2024)
Constrained Stein Variational Trajectory Optimization
by: Power, Thomas, et al.
Published: (2023)
by: Power, Thomas, et al.
Published: (2023)
Multimodal Safe Control for Human-Robot Interaction
by: Pandya, Ravi, et al.
Published: (2023)
by: Pandya, Ravi, et al.
Published: (2023)
Off-Policy Safe Reinforcement Learning with Constrained Optimistic Exploration
by: Li, Guopeng, et al.
Published: (2026)
by: Li, Guopeng, et al.
Published: (2026)
MePoly: Max Entropy Polynomial Policy Optimization
by: Liu, Hang, et al.
Published: (2026)
by: Liu, Hang, et al.
Published: (2026)
Action-Constrained Imitation Learning
by: Yeh, Chia-Han, et al.
Published: (2025)
by: Yeh, Chia-Han, et al.
Published: (2025)
SOMTP: Self-Supervised Learning-Based Optimizer for MPC-Based Safe Trajectory Planning Problems in Robotics
by: Liu, Yifan, et al.
Published: (2024)
by: Liu, Yifan, et al.
Published: (2024)
OPTIMA: Optimized Policy for Intelligent Multi-Agent Systems Enables Coordination-Aware Autonomous Vehicles
by: Du, Rui, et al.
Published: (2024)
by: Du, Rui, et al.
Published: (2024)
Adversarial Constrained Policy Optimization: Improving Constrained Reinforcement Learning by Adapting Budgets
by: Ma, Jianmina, et al.
Published: (2024)
by: Ma, Jianmina, et al.
Published: (2024)
CRAFT: Counterfactual-to-Interactive Reinforcement Fine-Tuning for Driving Policies
by: Chen, Keyu, et al.
Published: (2026)
by: Chen, Keyu, et al.
Published: (2026)
Efficient Policy Optimization in Robust Constrained MDPs with Iteration Complexity Guarantees
by: Ganguly, Sourav, et al.
Published: (2025)
by: Ganguly, Sourav, et al.
Published: (2025)
Certifying Robustness of Learning-Based Keypoint Detection and Pose Estimation Methods
by: Luo, Xusheng, et al.
Published: (2024)
by: Luo, Xusheng, et al.
Published: (2024)
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models
by: Lyu, Mingyang, et al.
Published: (2025)
by: Lyu, Mingyang, et al.
Published: (2025)
Query-Centric Diffusion Policy for Generalizable Robotic Assembly
by: Xu, Ziyi, et al.
Published: (2025)
by: Xu, Ziyi, et al.
Published: (2025)
FocalPolicy: Frequency-Optimized Chunking and Locally Anchored Flow Matching for Coherent Visuomotor Policy
by: He, Qian, et al.
Published: (2026)
by: He, Qian, et al.
Published: (2026)
ERPPO: Entropy Regularization-based Proximal Policy Optimization
by: Lee, Changha, et al.
Published: (2026)
by: Lee, Changha, et al.
Published: (2026)
Robust Adversarial Policy Optimization Under Dynamics Uncertainty
by: Kim, Mintae, et al.
Published: (2026)
by: Kim, Mintae, et al.
Published: (2026)
Multi-Agent Generative Adversarial Interactive Self-Imitation Learning for AUV Formation Control and Obstacle Avoidance
by: Fang, Zheng, et al.
Published: (2024)
by: Fang, Zheng, et al.
Published: (2024)
Similar Items
-
Learn With Imagination: Safe Set Guided State-wise Constrained Policy Optimization
by: Sun, Yifan, et al.
Published: (2023) -
Absolute Policy Optimization
by: Zhao, Weiye, et al.
Published: (2023) -
Absolute State-wise Constrained Policy Optimization: High-Probability State-wise Constraints Satisfaction
by: Zhao, Weiye, et al.
Published: (2024) -
GUARD: A Safe Reinforcement Learning Benchmark
by: Zhao, Weiye, et al.
Published: (2023) -
Physics-Aware Combinatorial Assembly Sequence Planning using Data-free Action Masking
by: Liu, Ruixuan, et al.
Published: (2024)