TW-CRL: Time-Weighted Contrastive Reward Learning for Efficient Inverse Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Yuxuan, Gao, Yicheng, Yang, Ning, Xia, Stephen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
IR$^3$: Contrastive Inverse Reinforcement Learning for Interpretable Detection and Mitigation of Reward Hacking
por: Beigi, Mohammad, et al.
Publicado: (2026)
por: Beigi, Mohammad, et al.
Publicado: (2026)
CRL-VLA: Continual Vision-Language-Action Learning
por: Zeng, Qixin, et al.
Publicado: (2026)
por: Zeng, Qixin, et al.
Publicado: (2026)
Revisiting Reinforcement Learning with Verifiable Rewards from a Contrastive Perspective
por: Zhang, Feng, et al.
Publicado: (2026)
por: Zhang, Feng, et al.
Publicado: (2026)
Adaptive Correlation-Weighted Intrinsic Rewards for Reinforcement Learning
por: Nguyen, Viet Bac, et al.
Publicado: (2026)
por: Nguyen, Viet Bac, et al.
Publicado: (2026)
Learning Rewards, Not Labels: Adversarial Inverse Reinforcement Learning for Machinery Fault Detection
por: Neupane, Dhiraj, et al.
Publicado: (2026)
por: Neupane, Dhiraj, et al.
Publicado: (2026)
Loss- and Reward-Weighting for Efficient Distributed Reinforcement Learning
por: Holen, Martin, et al.
Publicado: (2023)
por: Holen, Martin, et al.
Publicado: (2023)
Towards the Transferability of Rewards Recovered via Regularized Inverse Reinforcement Learning
por: Schlaginhaufen, Andreas, et al.
Publicado: (2024)
por: Schlaginhaufen, Andreas, et al.
Publicado: (2024)
Provably Efficient Exploration in Inverse Constrained Reinforcement Learning
por: Yue, Bo, et al.
Publicado: (2024)
por: Yue, Bo, et al.
Publicado: (2024)
Inverse Reinforcement Learning with Dynamic Reward Scaling for LLM Alignment
por: Cheng, Ruoxi, et al.
Publicado: (2025)
por: Cheng, Ruoxi, et al.
Publicado: (2025)
Efficient Reinforcement Learning in Probabilistic Reward Machines
por: Lin, Xiaofeng, et al.
Publicado: (2024)
por: Lin, Xiaofeng, et al.
Publicado: (2024)
Generalizing Behavior via Inverse Reinforcement Learning with Closed-Form Reward Centroids
por: Lazzati, Filippo, et al.
Publicado: (2025)
por: Lazzati, Filippo, et al.
Publicado: (2025)
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors
por: Ke, Jingyang, et al.
Publicado: (2025)
por: Ke, Jingyang, et al.
Publicado: (2025)
Enhancing Inverse Reinforcement Learning through Encoding Dynamic Information in Reward Shaping
por: Zhan, Simon Sinong, et al.
Publicado: (2024)
por: Zhan, Simon Sinong, et al.
Publicado: (2024)
Efficient $Q$-Learning and Actor-Critic Methods for Robust Average Reward Reinforcement Learning
por: Xu, Yang, et al.
Publicado: (2025)
por: Xu, Yang, et al.
Publicado: (2025)
K-Score: Kalman Filter as a Principled Alternative to Reward Normalization in Reinforcement Learning
por: Xia, Zixuan, et al.
Publicado: (2026)
por: Xia, Zixuan, et al.
Publicado: (2026)
Inverse Reinforcement Learning without an Optimal Demonstrator: A Feasible Reward Set Approach
por: Kim, Kihyun, et al.
Publicado: (2026)
por: Kim, Kihyun, et al.
Publicado: (2026)
Weighted Contrastive Learning for Anomaly-Aware Time-Series Forecasting
por: Ekstrand, Joel, et al.
Publicado: (2025)
por: Ekstrand, Joel, et al.
Publicado: (2025)
Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning
por: Ma, Haozhe, et al.
Publicado: (2024)
por: Ma, Haozhe, et al.
Publicado: (2024)
Latent-Space Contrastive Reinforcement Learning for Stable and Efficient LLM Reasoning
por: Shan, Lianlei, et al.
Publicado: (2026)
por: Shan, Lianlei, et al.
Publicado: (2026)
Hybrid Inverse Reinforcement Learning
por: Ren, Juntao, et al.
Publicado: (2024)
por: Ren, Juntao, et al.
Publicado: (2024)
Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards
por: Metcalf, Katherine, et al.
Publicado: (2024)
por: Metcalf, Katherine, et al.
Publicado: (2024)
Reward Models in Deep Reinforcement Learning: A Survey
por: Yu, Rui, et al.
Publicado: (2025)
por: Yu, Rui, et al.
Publicado: (2025)
Causal Information Prioritization for Efficient Reinforcement Learning
por: Cao, Hongye, et al.
Publicado: (2025)
por: Cao, Hongye, et al.
Publicado: (2025)
Kernel Density Bayesian Inverse Reinforcement Learning
por: Mandyam, Aishwarya, et al.
Publicado: (2023)
por: Mandyam, Aishwarya, et al.
Publicado: (2023)
Reinforcement Learning with Stochastic Reward Machines
por: Corazza, Jan, et al.
Publicado: (2025)
por: Corazza, Jan, et al.
Publicado: (2025)
Reinforcement Learning with Exogenous States and Rewards
por: Trimponias, George, et al.
Publicado: (2023)
por: Trimponias, George, et al.
Publicado: (2023)
Reinforcement Learning with Symbolic Reward Machines
por: Krug, Thomas, et al.
Publicado: (2026)
por: Krug, Thomas, et al.
Publicado: (2026)
Offline Reinforcement Learning with Imputed Rewards
por: Romeo, Carlo, et al.
Publicado: (2024)
por: Romeo, Carlo, et al.
Publicado: (2024)
Distillation Enhanced Time Series Forecasting Network with Momentum Contrastive Learning
por: Gao, Haozhi, et al.
Publicado: (2024)
por: Gao, Haozhi, et al.
Publicado: (2024)
Dynamic Adversarial Reinforcement Learning for Robust Multimodal Large Language Models
por: Bao, Yicheng, et al.
Publicado: (2026)
por: Bao, Yicheng, et al.
Publicado: (2026)
Contrastive Abstraction for Reinforcement Learning
por: Patil, Vihang, et al.
Publicado: (2024)
por: Patil, Vihang, et al.
Publicado: (2024)
Provably Efficient Reward Transfer in Reinforcement Learning with Discrete Markov Decision Processes
por: Vora, Kevin, et al.
Publicado: (2025)
por: Vora, Kevin, et al.
Publicado: (2025)
Recursive Deep Inverse Reinforcement Learning
por: Ghanem, Paul, et al.
Publicado: (2025)
por: Ghanem, Paul, et al.
Publicado: (2025)
Fast Rates for Inverse Reinforcement Learning
por: Schlaginhaufen, Andreas, et al.
Publicado: (2026)
por: Schlaginhaufen, Andreas, et al.
Publicado: (2026)
On the Effective Horizon of Inverse Reinforcement Learning
por: Xu, Yiqing, et al.
Publicado: (2023)
por: Xu, Yiqing, et al.
Publicado: (2023)
Environment Design for Inverse Reinforcement Learning
por: Buening, Thomas Kleine, et al.
Publicado: (2022)
por: Buening, Thomas Kleine, et al.
Publicado: (2022)
Inverse Reinforcement Learning With Constraint Recovery
por: Das, Nirjhar, et al.
Publicado: (2023)
por: Das, Nirjhar, et al.
Publicado: (2023)
DRTA: Dynamic Reward Scaling for Reinforcement Learning in Time Series Anomaly Detection
por: Golchin, Bahareh, et al.
Publicado: (2025)
por: Golchin, Bahareh, et al.
Publicado: (2025)
Inverse Design in Distributed Circuits Using Single-Step Reinforcement Learning
por: Li, Jiayu, et al.
Publicado: (2025)
por: Li, Jiayu, et al.
Publicado: (2025)
Inverse Delayed Reinforcement Learning
por: Zhan, Simon Sinong, et al.
Publicado: (2024)
por: Zhan, Simon Sinong, et al.
Publicado: (2024)
Ejemplares similares
-
IR$^3$: Contrastive Inverse Reinforcement Learning for Interpretable Detection and Mitigation of Reward Hacking
por: Beigi, Mohammad, et al.
Publicado: (2026) -
CRL-VLA: Continual Vision-Language-Action Learning
por: Zeng, Qixin, et al.
Publicado: (2026) -
Revisiting Reinforcement Learning with Verifiable Rewards from a Contrastive Perspective
por: Zhang, Feng, et al.
Publicado: (2026) -
Adaptive Correlation-Weighted Intrinsic Rewards for Reinforcement Learning
por: Nguyen, Viet Bac, et al.
Publicado: (2026) -
Learning Rewards, Not Labels: Adversarial Inverse Reinforcement Learning for Machinery Fault Detection
por: Neupane, Dhiraj, et al.
Publicado: (2026)