Can We Really Learn One Representation to Optimize All Rewards?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Chongyi, Jayanth, Royina Karegoudra, Eysenbach, Benjamin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can a MISL Fly? Analysis and Ingredients for Mutual Information Skill Learning
von: Zheng, Chongyi, et al.
Veröffentlicht: (2024)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2024)
Contrastive Difference Predictive Coding
von: Zheng, Chongyi, et al.
Veröffentlicht: (2023)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2023)
Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning
von: Venugopal, Aravind, et al.
Veröffentlicht: (2026)
von: Venugopal, Aravind, et al.
Veröffentlicht: (2026)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
von: Park, Seohong, et al.
Veröffentlicht: (2023)
von: Park, Seohong, et al.
Veröffentlicht: (2023)
Intention-Conditioned Flow Occupancy Models
von: Zheng, Chongyi, et al.
Veröffentlicht: (2025)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2025)
UpSkill: Mutual Information Skill Learning for Structured Response Diversity in LLMs
von: Shah, Devan, et al.
Veröffentlicht: (2026)
von: Shah, Devan, et al.
Veröffentlicht: (2026)
Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards
von: Mohamed, Faisal, et al.
Veröffentlicht: (2026)
von: Mohamed, Faisal, et al.
Veröffentlicht: (2026)
A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals
von: Liu, Grace, et al.
Veröffentlicht: (2024)
von: Liu, Grace, et al.
Veröffentlicht: (2024)
Value Flows
von: Dong, Perry, et al.
Veröffentlicht: (2025)
von: Dong, Perry, et al.
Veröffentlicht: (2025)
Neural Inertial Odometry from Lie Events
von: Jayanth, Royina Karegoudra, et al.
Veröffentlicht: (2025)
von: Jayanth, Royina Karegoudra, et al.
Veröffentlicht: (2025)
From Demonstrations to Rewards: Test-Time Prompt Optimization for VLM Reward Models
von: Gumbsch, Christian, et al.
Veröffentlicht: (2026)
von: Gumbsch, Christian, et al.
Veröffentlicht: (2026)
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
von: Zhang, Chen Bo Calvin, et al.
Veröffentlicht: (2024)
von: Zhang, Chen Bo Calvin, et al.
Veröffentlicht: (2024)
Can We Detect Failures Without Failure Data? Uncertainty-Aware Runtime Failure Detection for Imitation Learning Policies
von: Xu, Chen, et al.
Veröffentlicht: (2025)
von: Xu, Chen, et al.
Veröffentlicht: (2025)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
von: Ishihara, Yu, et al.
Veröffentlicht: (2025)
von: Ishihara, Yu, et al.
Veröffentlicht: (2025)
Generalized Animal Imitator: Agile Locomotion with Versatile Motion Prior
von: Yang, Ruihan, et al.
Veröffentlicht: (2023)
von: Yang, Ruihan, et al.
Veröffentlicht: (2023)
Diffusion-Reward Adversarial Imitation Learning
von: Lai, Chun-Mao, et al.
Veröffentlicht: (2024)
von: Lai, Chun-Mao, et al.
Veröffentlicht: (2024)
Reward-Punishment Reinforcement Learning with Maximum Entropy
von: Wang, Jiexin, et al.
Veröffentlicht: (2024)
von: Wang, Jiexin, et al.
Veröffentlicht: (2024)
TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal Distance
von: Liu, Yuyang, et al.
Veröffentlicht: (2025)
von: Liu, Yuyang, et al.
Veröffentlicht: (2025)
Assigning Credit with Partial Reward Decoupling in Multi-Agent Proximal Policy Optimization
von: Kapoor, Aditya, et al.
Veröffentlicht: (2024)
von: Kapoor, Aditya, et al.
Veröffentlicht: (2024)
Subwords as Skills: Tokenization for Sparse-Reward Reinforcement Learning
von: Yunis, David, et al.
Veröffentlicht: (2023)
von: Yunis, David, et al.
Veröffentlicht: (2023)
Adaptive Querying for Reward Learning from Human Feedback
von: Anand, Yashwanthi, et al.
Veröffentlicht: (2024)
von: Anand, Yashwanthi, et al.
Veröffentlicht: (2024)
Residual Reward Models for Preference-based Reinforcement Learning
von: Cao, Chenyang, et al.
Veröffentlicht: (2025)
von: Cao, Chenyang, et al.
Veröffentlicht: (2025)
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
von: Diaz-Bone, Leander, et al.
Veröffentlicht: (2025)
von: Diaz-Bone, Leander, et al.
Veröffentlicht: (2025)
Robot Policy Learning with Temporal Optimal Transport Reward
von: Fu, Yuwei, et al.
Veröffentlicht: (2024)
von: Fu, Yuwei, et al.
Veröffentlicht: (2024)
Enabling Option Learning in Sparse Rewards with Hindsight Experience Replay
von: Romio, Gabriel, et al.
Veröffentlicht: (2026)
von: Romio, Gabriel, et al.
Veröffentlicht: (2026)
Batch Active Learning of Reward Functions from Human Preferences
von: Bıyık, Erdem, et al.
Veröffentlicht: (2024)
von: Bıyık, Erdem, et al.
Veröffentlicht: (2024)
DrS: Learning Reusable Dense Rewards for Multi-Stage Tasks
von: Mu, Tongzhou, et al.
Veröffentlicht: (2024)
von: Mu, Tongzhou, et al.
Veröffentlicht: (2024)
CaRL: Learning Scalable Planning Policies with Simple Rewards
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2025)
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2025)
A Generalized Acquisition Function for Preference-based Reward Learning
von: Ellis, Evan, et al.
Veröffentlicht: (2024)
von: Ellis, Evan, et al.
Veröffentlicht: (2024)
Reward Learning from Suboptimal Demonstrations with Applications in Surgical Electrocautery
von: Karimi, Zohre, et al.
Veröffentlicht: (2024)
von: Karimi, Zohre, et al.
Veröffentlicht: (2024)
Stabilizing Contrastive RL: Techniques for Robotic Goal Reaching from Offline Data
von: Zheng, Chongyi, et al.
Veröffentlicht: (2023)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2023)
Causally Robust Reward Learning from Reason-Augmented Preference Feedback
von: Hwang, Minjune, et al.
Veröffentlicht: (2026)
von: Hwang, Minjune, et al.
Veröffentlicht: (2026)
A Review of Reward Functions for Reinforcement Learning in the context of Autonomous Driving
von: Abouelazm, Ahmed, et al.
Veröffentlicht: (2024)
von: Abouelazm, Ahmed, et al.
Veröffentlicht: (2024)
Learning to Recover: Dynamic Reward Shaping with Wheel-Leg Coordination for Fallen Robots
von: Deng, Boyuan, et al.
Veröffentlicht: (2025)
von: Deng, Boyuan, et al.
Veröffentlicht: (2025)
DreamSmooth: Improving Model-based Reinforcement Learning via Reward Smoothing
von: Lee, Vint, et al.
Veröffentlicht: (2023)
von: Lee, Vint, et al.
Veröffentlicht: (2023)
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation
von: Guo, Yihong, et al.
Veröffentlicht: (2024)
von: Guo, Yihong, et al.
Veröffentlicht: (2024)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
von: Patel, Bhrij, et al.
Veröffentlicht: (2023)
von: Patel, Bhrij, et al.
Veröffentlicht: (2023)
Beyond Scalar Rewards: Distributional Reinforcement Learning with Preordered Objectives for Safe and Reliable Autonomous Driving
von: Abouelazm, Ahmed, et al.
Veröffentlicht: (2026)
von: Abouelazm, Ahmed, et al.
Veröffentlicht: (2026)
RobotKeyframing: Learning Locomotion with High-Level Objectives via Mixture of Dense and Sparse Rewards
von: Zargarbashi, Fatemeh, et al.
Veröffentlicht: (2024)
von: Zargarbashi, Fatemeh, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Can a MISL Fly? Analysis and Ingredients for Mutual Information Skill Learning
von: Zheng, Chongyi, et al.
Veröffentlicht: (2024) -
Contrastive Difference Predictive Coding
von: Zheng, Chongyi, et al.
Veröffentlicht: (2023) -
Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making
von: Myers, Vivek, et al.
Veröffentlicht: (2024) -
Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning
von: Venugopal, Aravind, et al.
Veröffentlicht: (2026) -
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
von: Park, Seohong, et al.
Veröffentlicht: (2023)