TW-CRL: Time-Weighted Contrastive Reward Learning for Efficient Inverse Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yuxuan, Gao, Yicheng, Yang, Ning, Xia, Stephen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
IR$^3$: Contrastive Inverse Reinforcement Learning for Interpretable Detection and Mitigation of Reward Hacking
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026)
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026)
CRL-VLA: Continual Vision-Language-Action Learning
von: Zeng, Qixin, et al.
Veröffentlicht: (2026)
von: Zeng, Qixin, et al.
Veröffentlicht: (2026)
Revisiting Reinforcement Learning with Verifiable Rewards from a Contrastive Perspective
von: Zhang, Feng, et al.
Veröffentlicht: (2026)
von: Zhang, Feng, et al.
Veröffentlicht: (2026)
Adaptive Correlation-Weighted Intrinsic Rewards for Reinforcement Learning
von: Nguyen, Viet Bac, et al.
Veröffentlicht: (2026)
von: Nguyen, Viet Bac, et al.
Veröffentlicht: (2026)
Learning Rewards, Not Labels: Adversarial Inverse Reinforcement Learning for Machinery Fault Detection
von: Neupane, Dhiraj, et al.
Veröffentlicht: (2026)
von: Neupane, Dhiraj, et al.
Veröffentlicht: (2026)
Loss- and Reward-Weighting for Efficient Distributed Reinforcement Learning
von: Holen, Martin, et al.
Veröffentlicht: (2023)
von: Holen, Martin, et al.
Veröffentlicht: (2023)
Towards the Transferability of Rewards Recovered via Regularized Inverse Reinforcement Learning
von: Schlaginhaufen, Andreas, et al.
Veröffentlicht: (2024)
von: Schlaginhaufen, Andreas, et al.
Veröffentlicht: (2024)
Provably Efficient Exploration in Inverse Constrained Reinforcement Learning
von: Yue, Bo, et al.
Veröffentlicht: (2024)
von: Yue, Bo, et al.
Veröffentlicht: (2024)
Inverse Reinforcement Learning with Dynamic Reward Scaling for LLM Alignment
von: Cheng, Ruoxi, et al.
Veröffentlicht: (2025)
von: Cheng, Ruoxi, et al.
Veröffentlicht: (2025)
Efficient Reinforcement Learning in Probabilistic Reward Machines
von: Lin, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Lin, Xiaofeng, et al.
Veröffentlicht: (2024)
Generalizing Behavior via Inverse Reinforcement Learning with Closed-Form Reward Centroids
von: Lazzati, Filippo, et al.
Veröffentlicht: (2025)
von: Lazzati, Filippo, et al.
Veröffentlicht: (2025)
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors
von: Ke, Jingyang, et al.
Veröffentlicht: (2025)
von: Ke, Jingyang, et al.
Veröffentlicht: (2025)
Enhancing Inverse Reinforcement Learning through Encoding Dynamic Information in Reward Shaping
von: Zhan, Simon Sinong, et al.
Veröffentlicht: (2024)
von: Zhan, Simon Sinong, et al.
Veröffentlicht: (2024)
Efficient $Q$-Learning and Actor-Critic Methods for Robust Average Reward Reinforcement Learning
von: Xu, Yang, et al.
Veröffentlicht: (2025)
von: Xu, Yang, et al.
Veröffentlicht: (2025)
K-Score: Kalman Filter as a Principled Alternative to Reward Normalization in Reinforcement Learning
von: Xia, Zixuan, et al.
Veröffentlicht: (2026)
von: Xia, Zixuan, et al.
Veröffentlicht: (2026)
Inverse Reinforcement Learning without an Optimal Demonstrator: A Feasible Reward Set Approach
von: Kim, Kihyun, et al.
Veröffentlicht: (2026)
von: Kim, Kihyun, et al.
Veröffentlicht: (2026)
Weighted Contrastive Learning for Anomaly-Aware Time-Series Forecasting
von: Ekstrand, Joel, et al.
Veröffentlicht: (2025)
von: Ekstrand, Joel, et al.
Veröffentlicht: (2025)
Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning
von: Ma, Haozhe, et al.
Veröffentlicht: (2024)
von: Ma, Haozhe, et al.
Veröffentlicht: (2024)
Latent-Space Contrastive Reinforcement Learning for Stable and Efficient LLM Reasoning
von: Shan, Lianlei, et al.
Veröffentlicht: (2026)
von: Shan, Lianlei, et al.
Veröffentlicht: (2026)
Hybrid Inverse Reinforcement Learning
von: Ren, Juntao, et al.
Veröffentlicht: (2024)
von: Ren, Juntao, et al.
Veröffentlicht: (2024)
Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards
von: Metcalf, Katherine, et al.
Veröffentlicht: (2024)
von: Metcalf, Katherine, et al.
Veröffentlicht: (2024)
Reward Models in Deep Reinforcement Learning: A Survey
von: Yu, Rui, et al.
Veröffentlicht: (2025)
von: Yu, Rui, et al.
Veröffentlicht: (2025)
Causal Information Prioritization for Efficient Reinforcement Learning
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
Kernel Density Bayesian Inverse Reinforcement Learning
von: Mandyam, Aishwarya, et al.
Veröffentlicht: (2023)
von: Mandyam, Aishwarya, et al.
Veröffentlicht: (2023)
Reinforcement Learning with Stochastic Reward Machines
von: Corazza, Jan, et al.
Veröffentlicht: (2025)
von: Corazza, Jan, et al.
Veröffentlicht: (2025)
Reinforcement Learning with Exogenous States and Rewards
von: Trimponias, George, et al.
Veröffentlicht: (2023)
von: Trimponias, George, et al.
Veröffentlicht: (2023)
Reinforcement Learning with Symbolic Reward Machines
von: Krug, Thomas, et al.
Veröffentlicht: (2026)
von: Krug, Thomas, et al.
Veröffentlicht: (2026)
Offline Reinforcement Learning with Imputed Rewards
von: Romeo, Carlo, et al.
Veröffentlicht: (2024)
von: Romeo, Carlo, et al.
Veröffentlicht: (2024)
Distillation Enhanced Time Series Forecasting Network with Momentum Contrastive Learning
von: Gao, Haozhi, et al.
Veröffentlicht: (2024)
von: Gao, Haozhi, et al.
Veröffentlicht: (2024)
Dynamic Adversarial Reinforcement Learning for Robust Multimodal Large Language Models
von: Bao, Yicheng, et al.
Veröffentlicht: (2026)
von: Bao, Yicheng, et al.
Veröffentlicht: (2026)
Contrastive Abstraction for Reinforcement Learning
von: Patil, Vihang, et al.
Veröffentlicht: (2024)
von: Patil, Vihang, et al.
Veröffentlicht: (2024)
Provably Efficient Reward Transfer in Reinforcement Learning with Discrete Markov Decision Processes
von: Vora, Kevin, et al.
Veröffentlicht: (2025)
von: Vora, Kevin, et al.
Veröffentlicht: (2025)
Recursive Deep Inverse Reinforcement Learning
von: Ghanem, Paul, et al.
Veröffentlicht: (2025)
von: Ghanem, Paul, et al.
Veröffentlicht: (2025)
Fast Rates for Inverse Reinforcement Learning
von: Schlaginhaufen, Andreas, et al.
Veröffentlicht: (2026)
von: Schlaginhaufen, Andreas, et al.
Veröffentlicht: (2026)
On the Effective Horizon of Inverse Reinforcement Learning
von: Xu, Yiqing, et al.
Veröffentlicht: (2023)
von: Xu, Yiqing, et al.
Veröffentlicht: (2023)
Environment Design for Inverse Reinforcement Learning
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2022)
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2022)
Inverse Reinforcement Learning With Constraint Recovery
von: Das, Nirjhar, et al.
Veröffentlicht: (2023)
von: Das, Nirjhar, et al.
Veröffentlicht: (2023)
DRTA: Dynamic Reward Scaling for Reinforcement Learning in Time Series Anomaly Detection
von: Golchin, Bahareh, et al.
Veröffentlicht: (2025)
von: Golchin, Bahareh, et al.
Veröffentlicht: (2025)
Inverse Design in Distributed Circuits Using Single-Step Reinforcement Learning
von: Li, Jiayu, et al.
Veröffentlicht: (2025)
von: Li, Jiayu, et al.
Veröffentlicht: (2025)
Inverse Delayed Reinforcement Learning
von: Zhan, Simon Sinong, et al.
Veröffentlicht: (2024)
von: Zhan, Simon Sinong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
IR$^3$: Contrastive Inverse Reinforcement Learning for Interpretable Detection and Mitigation of Reward Hacking
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026) -
CRL-VLA: Continual Vision-Language-Action Learning
von: Zeng, Qixin, et al.
Veröffentlicht: (2026) -
Revisiting Reinforcement Learning with Verifiable Rewards from a Contrastive Perspective
von: Zhang, Feng, et al.
Veröffentlicht: (2026) -
Adaptive Correlation-Weighted Intrinsic Rewards for Reinforcement Learning
von: Nguyen, Viet Bac, et al.
Veröffentlicht: (2026) -
Learning Rewards, Not Labels: Adversarial Inverse Reinforcement Learning for Machinery Fault Detection
von: Neupane, Dhiraj, et al.
Veröffentlicht: (2026)