Uncertainty-aware Reward Design Process
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Yang, Zhou, Xiaolu, Ding, Bosong, Xin, Miao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning secondary tool affordances of human partners using iCub robot's egocentric data
von: Ding, Bosong, et al.
Veröffentlicht: (2024)
von: Ding, Bosong, et al.
Veröffentlicht: (2024)
Imitation of human motion achieves natural head movements for humanoid robots in an active-speaker detection task
von: Ding, Bosong, et al.
Veröffentlicht: (2024)
von: Ding, Bosong, et al.
Veröffentlicht: (2024)
Reward Redistribution via Gaussian Process Likelihood Estimation
von: Xiao, Minheng, et al.
Veröffentlicht: (2025)
von: Xiao, Minheng, et al.
Veröffentlicht: (2025)
CUQDS: Conformal Uncertainty Quantification under Distribution Shift for Trajectory Prediction
von: Huang, Huiqun, et al.
Veröffentlicht: (2024)
von: Huang, Huiqun, et al.
Veröffentlicht: (2024)
LORD: Large Models based Opposite Reward Design for Autonomous Driving
von: Ye, Xin, et al.
Veröffentlicht: (2024)
von: Ye, Xin, et al.
Veröffentlicht: (2024)
Momentum Based Reward Design for Low Emission Traffic Signal Control
von: Mundane, Chinmay, et al.
Veröffentlicht: (2026)
von: Mundane, Chinmay, et al.
Veröffentlicht: (2026)
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
von: Zhang, Chen Bo Calvin, et al.
Veröffentlicht: (2024)
von: Zhang, Chen Bo Calvin, et al.
Veröffentlicht: (2024)
ReLAM: Learning Anticipation Model for Rewarding Visual Robotic Manipulation
von: Tang, Nan, et al.
Veröffentlicht: (2025)
von: Tang, Nan, et al.
Veröffentlicht: (2025)
ELEMENTAL: Interactive Learning from Demonstrations and Vision-Language Models for Reward Design in Robotics
von: Chen, Letian, et al.
Veröffentlicht: (2024)
von: Chen, Letian, et al.
Veröffentlicht: (2024)
Uncertainty-aware Latent Safety Filters for Avoiding Out-of-Distribution Failures
von: Seo, Junwon, et al.
Veröffentlicht: (2025)
von: Seo, Junwon, et al.
Veröffentlicht: (2025)
Rewarding DINO: Predicting Dense Rewards with Vision Foundation Models
von: Krack, Pierre, et al.
Veröffentlicht: (2026)
von: Krack, Pierre, et al.
Veröffentlicht: (2026)
The Dark Side of Rich Rewards: Understanding and Mitigating Noise in VLM Rewards
von: Huang, Sukai, et al.
Veröffentlicht: (2024)
von: Huang, Sukai, et al.
Veröffentlicht: (2024)
STRIDE: Automating Reward Design, Deep Reinforcement Learning Training and Feedback Optimization in Humanoid Robotics Locomotion
von: Wu, Zhenwei, et al.
Veröffentlicht: (2025)
von: Wu, Zhenwei, et al.
Veröffentlicht: (2025)
TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal Distance
von: Liu, Yuyang, et al.
Veröffentlicht: (2025)
von: Liu, Yuyang, et al.
Veröffentlicht: (2025)
ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation
von: Feng, Youhe, et al.
Veröffentlicht: (2026)
von: Feng, Youhe, et al.
Veröffentlicht: (2026)
Dual-Granularity Contrastive Reward via Generated Episodic Guidance for Efficient Embodied RL
von: Liu, Xin, et al.
Veröffentlicht: (2026)
von: Liu, Xin, et al.
Veröffentlicht: (2026)
Reward Machine Inference for Robotic Manipulation
von: Baert, Mattijs, et al.
Veröffentlicht: (2024)
von: Baert, Mattijs, et al.
Veröffentlicht: (2024)
Safety-aware Causal Representation for Trustworthy Offline Reinforcement Learning in Autonomous Driving
von: Lin, Haohong, et al.
Veröffentlicht: (2023)
von: Lin, Haohong, et al.
Veröffentlicht: (2023)
Curriculum Reinforcement Learning for Complex Reward Functions
von: Freitag, Kilian, et al.
Veröffentlicht: (2024)
von: Freitag, Kilian, et al.
Veröffentlicht: (2024)
On-Robot Reinforcement Learning with Goal-Contrastive Rewards
von: Biza, Ondrej, et al.
Veröffentlicht: (2024)
von: Biza, Ondrej, et al.
Veröffentlicht: (2024)
Eureka: Human-Level Reward Design via Coding Large Language Models
von: Ma, Yecheng Jason, et al.
Veröffentlicht: (2023)
von: Ma, Yecheng Jason, et al.
Veröffentlicht: (2023)
Revisiting Sparse Rewards for Goal-Reaching Reinforcement Learning
von: Vasan, Gautham, et al.
Veröffentlicht: (2024)
von: Vasan, Gautham, et al.
Veröffentlicht: (2024)
Skill-aware Mutual Information Optimisation for Generalisation in Reinforcement Learning
von: Yu, Xuehui, et al.
Veröffentlicht: (2024)
von: Yu, Xuehui, et al.
Veröffentlicht: (2024)
Generalization in Deep Reinforcement Learning for Robotic Navigation by Reward Shaping
von: Miranda, Victor R. F., et al.
Veröffentlicht: (2022)
von: Miranda, Victor R. F., et al.
Veröffentlicht: (2022)
Text2Reward: Reward Shaping with Language Models for Reinforcement Learning
von: Xie, Tianbao, et al.
Veröffentlicht: (2023)
von: Xie, Tianbao, et al.
Veröffentlicht: (2023)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
von: Ishihara, Yu, et al.
Veröffentlicht: (2025)
von: Ishihara, Yu, et al.
Veröffentlicht: (2025)
Taxonomy-aware Dynamic Motion Generation on Hyperbolic Manifolds
von: Augenstein, Luis, et al.
Veröffentlicht: (2025)
von: Augenstein, Luis, et al.
Veröffentlicht: (2025)
Reward Prediction Error Prioritisation in Experience Replay: The RPE-PER Method
von: Yamani, Hoda, et al.
Veröffentlicht: (2025)
von: Yamani, Hoda, et al.
Veröffentlicht: (2025)
Adaptive Teaching in Heterogeneous Agents: Balancing Surprise in Sparse Reward Scenarios
von: Clark, Emma, et al.
Veröffentlicht: (2024)
von: Clark, Emma, et al.
Veröffentlicht: (2024)
TopoNav: Topological Navigation for Efficient Exploration in Sparse Reward Environments
von: Hossain, Jumman, et al.
Veröffentlicht: (2024)
von: Hossain, Jumman, et al.
Veröffentlicht: (2024)
Bridging the Human to Robot Dexterity Gap through Object-Oriented Rewards
von: Guzey, Irmak, et al.
Veröffentlicht: (2024)
von: Guzey, Irmak, et al.
Veröffentlicht: (2024)
TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics
von: Chen, Shirui, et al.
Veröffentlicht: (2026)
von: Chen, Shirui, et al.
Veröffentlicht: (2026)
Uncertainty Quantification Metrics for Deep Regression
von: Lind, Simon Kristoffersson, et al.
Veröffentlicht: (2024)
von: Lind, Simon Kristoffersson, et al.
Veröffentlicht: (2024)
Geometry-aware RL for Manipulation of Varying Shapes and Deformable Objects
von: Hoang, Tai, et al.
Veröffentlicht: (2025)
von: Hoang, Tai, et al.
Veröffentlicht: (2025)
Efficient Language-instructed Skill Acquisition via Reward-Policy Co-Evolution
von: Huang, Changxin, et al.
Veröffentlicht: (2024)
von: Huang, Changxin, et al.
Veröffentlicht: (2024)
Learning Emergent Gaits with Decentralized Phase Oscillators: on the role of Observations, Rewards, and Feedback
von: Zhang, Jenny, et al.
Veröffentlicht: (2024)
von: Zhang, Jenny, et al.
Veröffentlicht: (2024)
Average-Reward Maximum Entropy Reinforcement Learning for Underactuated Double Pendulum Tasks
von: Choe, Jean Seong Bjorn, et al.
Veröffentlicht: (2024)
von: Choe, Jean Seong Bjorn, et al.
Veröffentlicht: (2024)
DexSim2Real: Foundation Model-Guided Sim-to-Real Transfer for Generalizable Dexterous Manipulation
von: Zeng, Zijian, et al.
Veröffentlicht: (2026)
von: Zeng, Zijian, et al.
Veröffentlicht: (2026)
Neural-Network-Driven Reward Prediction as a Heuristic: Advancing Q-Learning for Mobile Robot Path Planning
von: Ji, Yiming, et al.
Veröffentlicht: (2024)
von: Ji, Yiming, et al.
Veröffentlicht: (2024)
From Demonstrations to Rewards: Test-Time Prompt Optimization for VLM Reward Models
von: Gumbsch, Christian, et al.
Veröffentlicht: (2026)
von: Gumbsch, Christian, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Learning secondary tool affordances of human partners using iCub robot's egocentric data
von: Ding, Bosong, et al.
Veröffentlicht: (2024) -
Imitation of human motion achieves natural head movements for humanoid robots in an active-speaker detection task
von: Ding, Bosong, et al.
Veröffentlicht: (2024) -
Reward Redistribution via Gaussian Process Likelihood Estimation
von: Xiao, Minheng, et al.
Veröffentlicht: (2025) -
CUQDS: Conformal Uncertainty Quantification under Distribution Shift for Trajectory Prediction
von: Huang, Huiqun, et al.
Veröffentlicht: (2024) -
LORD: Large Models based Opposite Reward Design for Autonomous Driving
von: Ye, Xin, et al.
Veröffentlicht: (2024)