Subwords as Skills: Tokenization for Sparse-Reward Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yunis, David, Jung, Justin, Dai, Falcon, Walter, Matthew |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
von: Diaz-Bone, Leander, et al.
Veröffentlicht: (2025)
von: Diaz-Bone, Leander, et al.
Veröffentlicht: (2025)
Offline Reinforcement Learning with Discrete Diffusion Skills
von: Qiao, RuiXi, et al.
Veröffentlicht: (2025)
von: Qiao, RuiXi, et al.
Veröffentlicht: (2025)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
von: Ishihara, Yu, et al.
Veröffentlicht: (2025)
von: Ishihara, Yu, et al.
Veröffentlicht: (2025)
Enabling Option Learning in Sparse Rewards with Hindsight Experience Replay
von: Romio, Gabriel, et al.
Veröffentlicht: (2026)
von: Romio, Gabriel, et al.
Veröffentlicht: (2026)
Reward-Punishment Reinforcement Learning with Maximum Entropy
von: Wang, Jiexin, et al.
Veröffentlicht: (2024)
von: Wang, Jiexin, et al.
Veröffentlicht: (2024)
Residual Reward Models for Preference-based Reinforcement Learning
von: Cao, Chenyang, et al.
Veröffentlicht: (2025)
von: Cao, Chenyang, et al.
Veröffentlicht: (2025)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
von: Patel, Bhrij, et al.
Veröffentlicht: (2023)
von: Patel, Bhrij, et al.
Veröffentlicht: (2023)
ReinforceGen: Hybrid Skill Policies with Automated Data Generation and Reinforcement Learning
von: Zhou, Zihan, et al.
Veröffentlicht: (2025)
von: Zhou, Zihan, et al.
Veröffentlicht: (2025)
TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics
von: Chen, Shirui, et al.
Veröffentlicht: (2026)
von: Chen, Shirui, et al.
Veröffentlicht: (2026)
Skill-aware Mutual Information Optimisation for Generalisation in Reinforcement Learning
von: Yu, Xuehui, et al.
Veröffentlicht: (2024)
von: Yu, Xuehui, et al.
Veröffentlicht: (2024)
A Review of Reward Functions for Reinforcement Learning in the context of Autonomous Driving
von: Abouelazm, Ahmed, et al.
Veröffentlicht: (2024)
von: Abouelazm, Ahmed, et al.
Veröffentlicht: (2024)
Learning Agile Soccer Skills for a Bipedal Robot with Deep Reinforcement Learning
von: Haarnoja, Tuomas, et al.
Veröffentlicht: (2023)
von: Haarnoja, Tuomas, et al.
Veröffentlicht: (2023)
RobotKeyframing: Learning Locomotion with High-Level Objectives via Mixture of Dense and Sparse Rewards
von: Zargarbashi, Fatemeh, et al.
Veröffentlicht: (2024)
von: Zargarbashi, Fatemeh, et al.
Veröffentlicht: (2024)
DreamSmooth: Improving Model-based Reinforcement Learning via Reward Smoothing
von: Lee, Vint, et al.
Veröffentlicht: (2023)
von: Lee, Vint, et al.
Veröffentlicht: (2023)
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation
von: Guo, Yihong, et al.
Veröffentlicht: (2024)
von: Guo, Yihong, et al.
Veröffentlicht: (2024)
Text2Reward: Reward Shaping with Language Models for Reinforcement Learning
von: Xie, Tianbao, et al.
Veröffentlicht: (2023)
von: Xie, Tianbao, et al.
Veröffentlicht: (2023)
SLIM: Skill Learning with Multiple Critics
von: Emukpere, David, et al.
Veröffentlicht: (2024)
von: Emukpere, David, et al.
Veröffentlicht: (2024)
Developing Driving Strategies Efficiently: A Skill-Based Hierarchical Reinforcement Learning Approach
von: Gurses, Yigit, et al.
Veröffentlicht: (2023)
von: Gurses, Yigit, et al.
Veröffentlicht: (2023)
VendiRL: A Framework for Self-Supervised Reinforcement Learning of Diversely Diverse Skills
von: Lintunen, Erik M.
Veröffentlicht: (2025)
von: Lintunen, Erik M.
Veröffentlicht: (2025)
Beyond Scalar Rewards: Distributional Reinforcement Learning with Preordered Objectives for Safe and Reliable Autonomous Driving
von: Abouelazm, Ahmed, et al.
Veröffentlicht: (2026)
von: Abouelazm, Ahmed, et al.
Veröffentlicht: (2026)
Leveraging Analytic Gradients in Provably Safe Reinforcement Learning
von: Walter, Tim, et al.
Veröffentlicht: (2025)
von: Walter, Tim, et al.
Veröffentlicht: (2025)
Tactical Decision Making for Autonomous Trucks by Deep Reinforcement Learning with Total Cost of Operation Based Reward
von: Pathare, Deepthi, et al.
Veröffentlicht: (2024)
von: Pathare, Deepthi, et al.
Veröffentlicht: (2024)
Leveraging Sub-Optimal Data for Human-in-the-Loop Reinforcement Learning
von: Muslimani, Calarina, et al.
Veröffentlicht: (2024)
von: Muslimani, Calarina, et al.
Veröffentlicht: (2024)
Diffusion-Reward Adversarial Imitation Learning
von: Lai, Chun-Mao, et al.
Veröffentlicht: (2024)
von: Lai, Chun-Mao, et al.
Veröffentlicht: (2024)
Equivariant Action Sampling for Reinforcement Learning and Planning
von: Zhao, Linfeng, et al.
Veröffentlicht: (2024)
von: Zhao, Linfeng, et al.
Veröffentlicht: (2024)
A Clean Slate for Offline Reinforcement Learning
von: Jackson, Matthew Thomas, et al.
Veröffentlicht: (2025)
von: Jackson, Matthew Thomas, et al.
Veröffentlicht: (2025)
Handling Delay in Real-Time Reinforcement Learning
von: Anokhin, Ivan, et al.
Veröffentlicht: (2025)
von: Anokhin, Ivan, et al.
Veröffentlicht: (2025)
Learning Parameterized Skills from Demonstrations
von: Gupta, Vedant, et al.
Veröffentlicht: (2025)
von: Gupta, Vedant, et al.
Veröffentlicht: (2025)
Not Only Rewards But Also Constraints: Applications on Legged Robot Locomotion
von: Kim, Yunho, et al.
Veröffentlicht: (2023)
von: Kim, Yunho, et al.
Veröffentlicht: (2023)
Redistributing Rewards Across Time and Agents for Multi-Agent Reinforcement Learning
von: Kapoor, Aditya, et al.
Veröffentlicht: (2025)
von: Kapoor, Aditya, et al.
Veröffentlicht: (2025)
Exploiting Symmetry in Dynamics for Model-Based Reinforcement Learning with Asymmetric Rewards
von: Sonmez, Yasin, et al.
Veröffentlicht: (2024)
von: Sonmez, Yasin, et al.
Veröffentlicht: (2024)
TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal Distance
von: Liu, Yuyang, et al.
Veröffentlicht: (2025)
von: Liu, Yuyang, et al.
Veröffentlicht: (2025)
RIME: Robust Preference-based Reinforcement Learning with Noisy Preferences
von: Cheng, Jie, et al.
Veröffentlicht: (2024)
von: Cheng, Jie, et al.
Veröffentlicht: (2024)
Uncertainty-Based Smooth Policy Regularisation for Reinforcement Learning with Few Demonstrations
von: Zhu, Yujie, et al.
Veröffentlicht: (2025)
von: Zhu, Yujie, et al.
Veröffentlicht: (2025)
Dynamic Contrastive Skill Learning with State-Transition Based Skill Clustering and Dynamic Length Adjustment
von: Choi, Jinwoo, et al.
Veröffentlicht: (2025)
von: Choi, Jinwoo, et al.
Veröffentlicht: (2025)
Adaptive Reinforcement Learning for Unobservable Random Delays
von: Wikman, John, et al.
Veröffentlicht: (2025)
von: Wikman, John, et al.
Veröffentlicht: (2025)
Adaptive Querying for Reward Learning from Human Feedback
von: Anand, Yashwanthi, et al.
Veröffentlicht: (2024)
von: Anand, Yashwanthi, et al.
Veröffentlicht: (2024)
Robot Policy Learning with Temporal Optimal Transport Reward
von: Fu, Yuwei, et al.
Veröffentlicht: (2024)
von: Fu, Yuwei, et al.
Veröffentlicht: (2024)
Skill Generalization with Verbs
von: Ma, Rachel, et al.
Veröffentlicht: (2024)
von: Ma, Rachel, et al.
Veröffentlicht: (2024)
ViReSkill: Vision-Grounded Replanning with Skill Memory for LLM-Based Planning in Lifelong Robot Learning
von: Kagaya, Tomoyuki, et al.
Veröffentlicht: (2025)
von: Kagaya, Tomoyuki, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
von: Diaz-Bone, Leander, et al.
Veröffentlicht: (2025) -
Offline Reinforcement Learning with Discrete Diffusion Skills
von: Qiao, RuiXi, et al.
Veröffentlicht: (2025) -
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
von: Ishihara, Yu, et al.
Veröffentlicht: (2025) -
Enabling Option Learning in Sparse Rewards with Hindsight Experience Replay
von: Romio, Gabriel, et al.
Veröffentlicht: (2026) -
Reward-Punishment Reinforcement Learning with Maximum Entropy
von: Wang, Jiexin, et al.
Veröffentlicht: (2024)