Enhance Exploration in Safe Reinforcement Learning with Contrastive Representation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Doan, Duc Kien, Le, Bang Giang, Ta, Viet Cuong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Toward Finding Strong Pareto Optimal Policies in Multi-Agent Reinforcement Learning
by: Le, Bang Giang, et al.
Published: (2024)
by: Le, Bang Giang, et al.
Published: (2024)
Resolve Highway Conflict in Multi-Autonomous Vehicle Controls with Local State Attention
by: Ta, Xuan Duy, et al.
Published: (2025)
by: Ta, Xuan Duy, et al.
Published: (2025)
Explicit Credit Assignment through Local Rewards and Dependence Graphs in Multi-Agent Reinforcement Learning
by: Le, Bang Giang, et al.
Published: (2026)
by: Le, Bang Giang, et al.
Published: (2026)
Efficient Bilevel Optimization for Meta Label Correction in Noisy Label Learning
by: Nguyen, Ba Hoang Anh, et al.
Published: (2026)
by: Nguyen, Ba Hoang Anh, et al.
Published: (2026)
Balancing Knowledge Distillation for Imbalance Learning with Bilevel Optimization
by: Nguyen, Anh B. H., et al.
Published: (2026)
by: Nguyen, Anh B. H., et al.
Published: (2026)
Reinforcement Learning by Guided Safe Exploration
by: Yang, Qisong, et al.
Published: (2023)
by: Yang, Qisong, et al.
Published: (2023)
Virtual Fusion with Contrastive Learning for Single Sensor-based Activity Recognition
by: Nguyen, Duc-Anh, et al.
Published: (2023)
by: Nguyen, Duc-Anh, et al.
Published: (2023)
Variable-Agnostic Causal Exploration for Reinforcement Learning
by: Nguyen, Minh Hoang, et al.
Published: (2024)
by: Nguyen, Minh Hoang, et al.
Published: (2024)
Improving Graph Convolutional Networks with Transformer Layer in social-based items recommendation
by: Hoang, Thi Linh, et al.
Published: (2024)
by: Hoang, Thi Linh, et al.
Published: (2024)
Revisiting Safe Exploration in Safe Reinforcement learning
by: Eckel, David, et al.
Published: (2024)
by: Eckel, David, et al.
Published: (2024)
MIR: Efficient Exploration in Episodic Multi-Agent Reinforcement Learning via Mutual Intrinsic Reward
by: Chen, Kesheng, et al.
Published: (2025)
by: Chen, Kesheng, et al.
Published: (2025)
Safe-Support Q-Learning: Learning without Unsafe Exploration
by: Lim, Yeeun, et al.
Published: (2026)
by: Lim, Yeeun, et al.
Published: (2026)
Contrastive Representation for Data Filtering in Cross-Domain Offline Reinforcement Learning
by: Wen, Xiaoyu, et al.
Published: (2024)
by: Wen, Xiaoyu, et al.
Published: (2024)
Learning Structural Causal Models from Ordering: Identifiable Flow Models
by: Le, Minh Khoa, et al.
Published: (2024)
by: Le, Minh Khoa, et al.
Published: (2024)
Enhancing Chess Reinforcement Learning with Graph Representation
by: Rigaux, Tomas, et al.
Published: (2024)
by: Rigaux, Tomas, et al.
Published: (2024)
Learning Reconfigurable Representations for Multimodal Federated Learning with Missing Data
by: Nguyen, Duong M., et al.
Published: (2025)
by: Nguyen, Duong M., et al.
Published: (2025)
Flatness-aware Sequential Learning Generates Resilient Backdoors
by: Pham, Hoang, et al.
Published: (2024)
by: Pham, Hoang, et al.
Published: (2024)
Probabilistic Shielding for Safe Reinforcement Learning
by: Court, Edwin Hamel-De le, et al.
Published: (2025)
by: Court, Edwin Hamel-De le, et al.
Published: (2025)
Contrastive Abstraction for Reinforcement Learning
by: Patil, Vihang, et al.
Published: (2024)
by: Patil, Vihang, et al.
Published: (2024)
Sequence Diffusion Model for Temporal Link Prediction in Continuous-Time Dynamic Graph
by: Duc, Nguyen Minh, et al.
Published: (2026)
by: Duc, Nguyen Minh, et al.
Published: (2026)
Adaptive Correlation-Weighted Intrinsic Rewards for Reinforcement Learning
by: Nguyen, Viet Bac, et al.
Published: (2026)
by: Nguyen, Viet Bac, et al.
Published: (2026)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
by: Anisimov, Maksim, et al.
Published: (2026)
by: Anisimov, Maksim, et al.
Published: (2026)
Safe Flow Q-Learning: Offline Safe Reinforcement Learning with Reachability-Based Flow Policies
by: Tayal, Mumuksh, et al.
Published: (2026)
by: Tayal, Mumuksh, et al.
Published: (2026)
In-context Exploration-Exploitation for Reinforcement Learning
by: Dai, Zhenwen, et al.
Published: (2024)
by: Dai, Zhenwen, et al.
Published: (2024)
Satisficing Exploration for Deep Reinforcement Learning
by: Arumugam, Dilip, et al.
Published: (2024)
by: Arumugam, Dilip, et al.
Published: (2024)
Safe Reinforcement Learning with Learned Non-Markovian Safety Constraints
by: Low, Siow Meng, et al.
Published: (2024)
by: Low, Siow Meng, et al.
Published: (2024)
Contrastive Reasoning Alignment: Reinforcement Learning from Hidden Representations
by: Luo, Haozheng, et al.
Published: (2026)
by: Luo, Haozheng, et al.
Published: (2026)
Online Optimization for Offline Safe Reinforcement Learning
by: Chemingui, Yassine, et al.
Published: (2025)
by: Chemingui, Yassine, et al.
Published: (2025)
Contrastive Learning Via Equivariant Representation
by: Song, Sifan, et al.
Published: (2024)
by: Song, Sifan, et al.
Published: (2024)
Self-Reinforced Graph Contrastive Learning
by: Hsieh, Chou-Ying, et al.
Published: (2025)
by: Hsieh, Chou-Ying, et al.
Published: (2025)
Sampling-Based Safe Reinforcement Learning
by: Vignola, Luca, et al.
Published: (2026)
by: Vignola, Luca, et al.
Published: (2026)
Implicit Safe Set Algorithm for Provably Safe Reinforcement Learning
by: Zhao, Weiye, et al.
Published: (2024)
by: Zhao, Weiye, et al.
Published: (2024)
Divide and Refine: Enhancing Multimodal Representation and Explainability for Emotion Recognition in Conversation
by: Mai, Anh-Tuan, et al.
Published: (2026)
by: Mai, Anh-Tuan, et al.
Published: (2026)
Is Exploration or Optimization the Problem for Deep Reinforcement Learning?
by: Berseth, Glen
Published: (2025)
by: Berseth, Glen
Published: (2025)
Neighboring State-based Exploration for Reinforcement Learning
by: Li, Yu-Teng, et al.
Published: (2022)
by: Li, Yu-Teng, et al.
Published: (2022)
Exploration in Knowledge Transfer Utilizing Reinforcement Learning
by: Jedlička, Adam, et al.
Published: (2024)
by: Jedlička, Adam, et al.
Published: (2024)
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning
by: Guo, Weiran, et al.
Published: (2025)
by: Guo, Weiran, et al.
Published: (2025)
Safe Reinforcement Learning for Real-World Engine Control
by: Bedei, Julian, et al.
Published: (2025)
by: Bedei, Julian, et al.
Published: (2025)
Skill-based Safe Reinforcement Learning with Risk Planning
by: Zhang, Hanping, et al.
Published: (2025)
by: Zhang, Hanping, et al.
Published: (2025)
Offline Safe Reinforcement Learning Using Trajectory Classification
by: Gong, Ze, et al.
Published: (2024)
by: Gong, Ze, et al.
Published: (2024)
Similar Items
-
Toward Finding Strong Pareto Optimal Policies in Multi-Agent Reinforcement Learning
by: Le, Bang Giang, et al.
Published: (2024) -
Resolve Highway Conflict in Multi-Autonomous Vehicle Controls with Local State Attention
by: Ta, Xuan Duy, et al.
Published: (2025) -
Explicit Credit Assignment through Local Rewards and Dependence Graphs in Multi-Agent Reinforcement Learning
by: Le, Bang Giang, et al.
Published: (2026) -
Efficient Bilevel Optimization for Meta Label Correction in Noisy Label Learning
by: Nguyen, Ba Hoang Anh, et al.
Published: (2026) -
Balancing Knowledge Distillation for Imbalance Learning with Bilevel Optimization
by: Nguyen, Anh B. H., et al.
Published: (2026)