Constrained Policy Optimization via Sampling-Based Weight-Space Projection
Fuente:
arXiv
Saved in:
| Main Authors: | Cao, Shengfan, Borrelli, Francesco, Joa, Eunhyek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Simple Approach to Constraint-Aware Imitation Learning with Application to Autonomous Racing
by: Cao, Shengfan, et al.
Published: (2025)
by: Cao, Shengfan, et al.
Published: (2025)
Approximate solution of stochastic infinite horizon optimal control problems for constrained linear uncertain systems
by: Joa, Eunhyek, et al.
Published: (2024)
by: Joa, Eunhyek, et al.
Published: (2024)
Energy-Aware Lane Planning for Connected Electric Vehicles in Urban Traffic: Design and Vehicle-in-the-Loop Validation
by: Kim, Hansung, et al.
Published: (2025)
by: Kim, Hansung, et al.
Published: (2025)
State-wise Constrained Policy Optimization
by: Zhao, Weiye, et al.
Published: (2023)
by: Zhao, Weiye, et al.
Published: (2023)
Energy-efficient predictive control for connected, automated driving under localization uncertainty
by: Joa, Eunhyek, et al.
Published: (2024)
by: Joa, Eunhyek, et al.
Published: (2024)
Eco-driving under localization uncertainty for connected vehicles on Urban roads: Data-driven approach and Experiment verification
by: Joa, Eunhyek, et al.
Published: (2024)
by: Joa, Eunhyek, et al.
Published: (2024)
Scalable Multi-modal Model Predictive Control via Duality-based Interaction Predictions
by: Kim, Hansung, et al.
Published: (2024)
by: Kim, Hansung, et al.
Published: (2024)
Constrained Group Relative Policy Optimization
by: Girgis, Roger, et al.
Published: (2026)
by: Girgis, Roger, et al.
Published: (2026)
Contractive Diffusion Policies: Robust Action Diffusion via Contractive Score-Based Sampling with Differential Equations
by: Abyaneh, Amin, et al.
Published: (2026)
by: Abyaneh, Amin, et al.
Published: (2026)
SAC Flow: Sample-Efficient Reinforcement Learning of Flow-Based Policies via Velocity-Reparameterized Sequential Modeling
by: Zhang, Yixian, et al.
Published: (2025)
by: Zhang, Yixian, et al.
Published: (2025)
State-Conditional Adversarial Learning: An Off-Policy Visual Domain Transfer Method for End-to-End Imitation Learning
by: Liu, Yuxiang, et al.
Published: (2025)
by: Liu, Yuxiang, et al.
Published: (2025)
Diffusion Policy Policy Optimization
by: Ren, Allen Z., et al.
Published: (2024)
by: Ren, Allen Z., et al.
Published: (2024)
Constrained Stein Variational Trajectory Optimization
by: Power, Thomas, et al.
Published: (2023)
by: Power, Thomas, et al.
Published: (2023)
Enhancing Sample Efficiency and Exploration in Reinforcement Learning through the Integration of Diffusion Models and Proximal Policy Optimization
by: Gao, Tianci, et al.
Published: (2024)
by: Gao, Tianci, et al.
Published: (2024)
Off-Policy Safe Reinforcement Learning with Constrained Optimistic Exploration
by: Li, Guopeng, et al.
Published: (2026)
by: Li, Guopeng, et al.
Published: (2026)
Diffusion Policy through Conditional Proximal Policy Optimization
by: Liu, Ben, et al.
Published: (2026)
by: Liu, Ben, et al.
Published: (2026)
M3PO: Massively Multi-Task Model-Based Policy Optimization
by: Narendra, Aditya, et al.
Published: (2025)
by: Narendra, Aditya, et al.
Published: (2025)
Policy Adaptation via Language Optimization: Decomposing Tasks for Few-Shot Imitation
by: Myers, Vivek, et al.
Published: (2024)
by: Myers, Vivek, et al.
Published: (2024)
Dichotomous Diffusion Policy Optimization
by: Liang, Ruiming, et al.
Published: (2025)
by: Liang, Ruiming, et al.
Published: (2025)
Adversarial Constrained Policy Optimization: Improving Constrained Reinforcement Learning by Adapting Budgets
by: Ma, Jianmina, et al.
Published: (2024)
by: Ma, Jianmina, et al.
Published: (2024)
Hybrid Machine Learning Model with a Constrained Action Space for Trajectory Prediction
by: Fertig, Alexander, et al.
Published: (2025)
by: Fertig, Alexander, et al.
Published: (2025)
NLP Sampling: Combining MCMC and NLP Methods for Diverse Constrained Sampling
by: Toussaint, Marc, et al.
Published: (2024)
by: Toussaint, Marc, et al.
Published: (2024)
OGPO: Sample Efficient Full-Finetuning of Generative Control Policies
by: Patil, Sarvesh, et al.
Published: (2026)
by: Patil, Sarvesh, et al.
Published: (2026)
Contractive Dynamical Imitation Policies for Efficient Out-of-Sample Recovery
by: Abyaneh, Amin, et al.
Published: (2024)
by: Abyaneh, Amin, et al.
Published: (2024)
PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning
by: Yang, Shunpeng, et al.
Published: (2026)
by: Yang, Shunpeng, et al.
Published: (2026)
SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows
by: Yang, Chenyu, et al.
Published: (2026)
by: Yang, Chenyu, et al.
Published: (2026)
Efficient Policy Optimization in Robust Constrained MDPs with Iteration Complexity Guarantees
by: Ganguly, Sourav, et al.
Published: (2025)
by: Ganguly, Sourav, et al.
Published: (2025)
X-IL: Exploring the Design Space of Imitation Learning Policies
by: Jia, Xiaogang, et al.
Published: (2025)
by: Jia, Xiaogang, et al.
Published: (2025)
Steering Your Diffusion Policy with Latent Space Reinforcement Learning
by: Wagenmaker, Andrew, et al.
Published: (2025)
by: Wagenmaker, Andrew, et al.
Published: (2025)
IMLE Policy: Fast and Sample Efficient Visuomotor Policy Learning via Implicit Maximum Likelihood Estimation
by: Rana, Krishan, et al.
Published: (2025)
by: Rana, Krishan, et al.
Published: (2025)
The Lie We Tell: Correcting the Euclidean Fallacy in Vision Language Action Policies via Score Matching on Tangent Space
by: Chuang, Bing-Cheng, et al.
Published: (2026)
by: Chuang, Bing-Cheng, et al.
Published: (2026)
VEGA: Electric Vehicle Navigation Agent via Physics-Informed Neural Operator and Proximal Policy Optimization
by: Lim, Hansol, et al.
Published: (2025)
by: Lim, Hansol, et al.
Published: (2025)
COSBO: Conservative Offline Simulation-Based Policy Optimization
by: Kargar, Eshagh, et al.
Published: (2024)
by: Kargar, Eshagh, et al.
Published: (2024)
Using Temperature Sampling to Effectively Train Robot Learning Policies on Imbalanced Datasets
by: Patil, Basavasagar, et al.
Published: (2025)
by: Patil, Basavasagar, et al.
Published: (2025)
SATA: Safe and Adaptive Torque-Based Locomotion Policies Inspired by Animal Learning
by: Li, Peizhuo, et al.
Published: (2025)
by: Li, Peizhuo, et al.
Published: (2025)
SOE: Sample-Efficient Robot Policy Self-Improvement via On-Manifold Exploration
by: Jin, Yang, et al.
Published: (2025)
by: Jin, Yang, et al.
Published: (2025)
FocalPolicy: Frequency-Optimized Chunking and Locally Anchored Flow Matching for Coherent Visuomotor Policy
by: He, Qian, et al.
Published: (2026)
by: He, Qian, et al.
Published: (2026)
ERPPO: Entropy Regularization-based Proximal Policy Optimization
by: Lee, Changha, et al.
Published: (2026)
by: Lee, Changha, et al.
Published: (2026)
Robust Adversarial Policy Optimization Under Dynamics Uncertainty
by: Kim, Mintae, et al.
Published: (2026)
by: Kim, Mintae, et al.
Published: (2026)
MePoly: Max Entropy Polynomial Policy Optimization
by: Liu, Hang, et al.
Published: (2026)
by: Liu, Hang, et al.
Published: (2026)
Similar Items
-
A Simple Approach to Constraint-Aware Imitation Learning with Application to Autonomous Racing
by: Cao, Shengfan, et al.
Published: (2025) -
Approximate solution of stochastic infinite horizon optimal control problems for constrained linear uncertain systems
by: Joa, Eunhyek, et al.
Published: (2024) -
Energy-Aware Lane Planning for Connected Electric Vehicles in Urban Traffic: Design and Vehicle-in-the-Loop Validation
by: Kim, Hansung, et al.
Published: (2025) -
State-wise Constrained Policy Optimization
by: Zhao, Weiye, et al.
Published: (2023) -
Energy-efficient predictive control for connected, automated driving under localization uncertainty
by: Joa, Eunhyek, et al.
Published: (2024)