Feasibility Consistent Representation Learning for Safe Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Cen, Zhepeng, Yao, Yihang, Liu, Zuxin, Zhao, Ding |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Behavior Injection: Preparing Language Models for Reinforcement Learning
by: Cen, Zhepeng, et al.
Published: (2025)
by: Cen, Zhepeng, et al.
Published: (2025)
Constraint-Conditioned Policy Optimization for Versatile Safe Reinforcement Learning
by: Yao, Yihang, et al.
Published: (2023)
by: Yao, Yihang, et al.
Published: (2023)
Learning from Sparse Offline Datasets via Conservative Density Estimation
by: Cen, Zhepeng, et al.
Published: (2024)
by: Cen, Zhepeng, et al.
Published: (2024)
OASIS: Conditional Distribution Shaping for Offline Safe Reinforcement Learning
by: Yao, Yihang, et al.
Published: (2024)
by: Yao, Yihang, et al.
Published: (2024)
Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization
by: Yao, Yihang, et al.
Published: (2026)
by: Yao, Yihang, et al.
Published: (2026)
Safety-aware Causal Representation for Trustworthy Offline Reinforcement Learning in Autonomous Driving
by: Lin, Haohong, et al.
Published: (2023)
by: Lin, Haohong, et al.
Published: (2023)
Feasible Policy Iteration for Safe Reinforcement Learning
by: Yang, Yujie, et al.
Published: (2023)
by: Yang, Yujie, et al.
Published: (2023)
Extreme Value Policy Optimization for Safe Reinforcement Learning
by: Gao, Shiqing, et al.
Published: (2026)
by: Gao, Shiqing, et al.
Published: (2026)
Safe Offline Reinforcement Learning with Feasibility-Guided Diffusion Model
by: Zheng, Yinan, et al.
Published: (2024)
by: Zheng, Yinan, et al.
Published: (2024)
Tailored Primitive Initialization is the Secret Key to Reinforcement Learning
by: Yao, Yihang, et al.
Published: (2025)
by: Yao, Yihang, et al.
Published: (2025)
Your Language Model May Think Too Rigidly: Achieving Reasoning Consistency with Symmetry-Enhanced Training
by: Yao, Yihang, et al.
Published: (2025)
by: Yao, Yihang, et al.
Published: (2025)
Policy Bifurcation in Safe Reinforcement Learning
by: Zou, Wenjun, et al.
Published: (2024)
by: Zou, Wenjun, et al.
Published: (2024)
Enhance Exploration in Safe Reinforcement Learning with Contrastive Representation Learning
by: Doan, Duc Kien, et al.
Published: (2025)
by: Doan, Duc Kien, et al.
Published: (2025)
Implicit Safe Set Algorithm for Provably Safe Reinforcement Learning
by: Zhao, Weiye, et al.
Published: (2024)
by: Zhao, Weiye, et al.
Published: (2024)
TAIL: Task-specific Adapters for Imitation Learning with Large Pretrained Models
by: Liu, Zuxin, et al.
Published: (2023)
by: Liu, Zuxin, et al.
Published: (2023)
Bridging the Training-Inference Gap in LLMs by Leveraging Self-Generated Tokens
by: Cen, Zhepeng, et al.
Published: (2024)
by: Cen, Zhepeng, et al.
Published: (2024)
Controlling Underestimation Bias in Constrained Reinforcement Learning for Safe Exploration
by: Gao, Shiqing, et al.
Published: (2026)
by: Gao, Shiqing, et al.
Published: (2026)
Vulnerability Analysis of Safe Reinforcement Learning via Inverse Constrained Reinforcement Learning
by: Fan, Jialiang, et al.
Published: (2026)
by: Fan, Jialiang, et al.
Published: (2026)
Consistency Models as a Rich and Efficient Policy Class for Reinforcement Learning
by: Ding, Zihan, et al.
Published: (2023)
by: Ding, Zihan, et al.
Published: (2023)
Safe In-Context Reinforcement Learning
by: Moeini, Amir, et al.
Published: (2025)
by: Moeini, Amir, et al.
Published: (2025)
Counterfactually Safe Reinforcement Learning
by: Li, Jingyi, et al.
Published: (2026)
by: Li, Jingyi, et al.
Published: (2026)
GUARD: A Safe Reinforcement Learning Benchmark
by: Zhao, Weiye, et al.
Published: (2023)
by: Zhao, Weiye, et al.
Published: (2023)
Enhancing Efficiency of Safe Reinforcement Learning via Sample Manipulation
by: Gu, Shangding, et al.
Published: (2024)
by: Gu, Shangding, et al.
Published: (2024)
A CMDP-within-online framework for Meta-Safe Reinforcement Learning
by: Khattar, Vanshaj, et al.
Published: (2024)
by: Khattar, Vanshaj, et al.
Published: (2024)
Global Safe Sequential Learning via Efficient Knowledge Transfer
by: Li, Cen-You, et al.
Published: (2024)
by: Li, Cen-You, et al.
Published: (2024)
Consistent Prompting for Rehearsal-Free Continual Learning
by: Gao, Zhanxin, et al.
Published: (2024)
by: Gao, Zhanxin, et al.
Published: (2024)
Learning Future Representation with Synthetic Observations for Sample-efficient Reinforcement Learning
by: Liu, Xin, et al.
Published: (2024)
by: Liu, Xin, et al.
Published: (2024)
Revisiting Bisimulation Metric for Robust Representations in Reinforcement Learning
by: Zhang, Leiji, et al.
Published: (2025)
by: Zhang, Leiji, et al.
Published: (2025)
Safe Reinforcement Learning on the Constraint Manifold: Theory and Applications
by: Liu, Puze, et al.
Published: (2024)
by: Liu, Puze, et al.
Published: (2024)
Conditional Sequence Modeling for Safe Reinforcement Learning
by: Bai, Wensong, et al.
Published: (2026)
by: Bai, Wensong, et al.
Published: (2026)
On Feasible Rewards in Multi-Agent Inverse Reinforcement Learning
by: Freihaut, Till, et al.
Published: (2024)
by: Freihaut, Till, et al.
Published: (2024)
BECAUSE: Bilinear Causal Representation for Generalizable Offline Model-based Reinforcement Learning
by: Lin, Haohong, et al.
Published: (2024)
by: Lin, Haohong, et al.
Published: (2024)
Safety through Permissibility: Shield Construction for Fast and Safe Reinforcement Learning
by: Politowicz, Alexander, et al.
Published: (2024)
by: Politowicz, Alexander, et al.
Published: (2024)
CIMRL: Combining IMitation and Reinforcement Learning for Safe Autonomous Driving
by: Booher, Jonathan, et al.
Published: (2024)
by: Booher, Jonathan, et al.
Published: (2024)
LexiSafe: Offline Safe Reinforcement Learning with Lexicographic Safety-Reward Hierarchy
by: Yang, Hsin-Jung, et al.
Published: (2026)
by: Yang, Hsin-Jung, et al.
Published: (2026)
Reinforcement Learning by Guided Safe Exploration
by: Yang, Qisong, et al.
Published: (2023)
by: Yang, Qisong, et al.
Published: (2023)
Probabilistic Shielding for Safe Reinforcement Learning
by: Court, Edwin Hamel-De le, et al.
Published: (2025)
by: Court, Edwin Hamel-De le, et al.
Published: (2025)
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning
by: Guo, Weiran, et al.
Published: (2025)
by: Guo, Weiran, et al.
Published: (2025)
Adversarially Trained Weighted Actor-Critic for Safe Offline Reinforcement Learning
by: Wei, Honghao, et al.
Published: (2024)
by: Wei, Honghao, et al.
Published: (2024)
Bridging the gap between Learning-to-plan, Motion Primitives and Safe Reinforcement Learning
by: Kicki, Piotr, et al.
Published: (2024)
by: Kicki, Piotr, et al.
Published: (2024)
Similar Items
-
Behavior Injection: Preparing Language Models for Reinforcement Learning
by: Cen, Zhepeng, et al.
Published: (2025) -
Constraint-Conditioned Policy Optimization for Versatile Safe Reinforcement Learning
by: Yao, Yihang, et al.
Published: (2023) -
Learning from Sparse Offline Datasets via Conservative Density Estimation
by: Cen, Zhepeng, et al.
Published: (2024) -
OASIS: Conditional Distribution Shaping for Offline Safe Reinforcement Learning
by: Yao, Yihang, et al.
Published: (2024) -
Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization
by: Yao, Yihang, et al.
Published: (2026)