Position: Good Embodied Reward Models Need Bad Behavior Data
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tian, Ran, Wu, Yilin, Bajcsy, Andrea |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Foresight to Forethought: VLM-In-the-Loop Policy Steering via Latent Alignment
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment
von: Tian, Ran, et al.
Veröffentlicht: (2024)
von: Tian, Ran, et al.
Veröffentlicht: (2024)
Not All Errors Are Made Equal: A Regret Metric for Detecting System-level Trajectory Prediction Failures
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2024)
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2024)
When to Act, Ask, or Learn: Uncertainty-Aware Policy Steering
von: Yuan, Jessie, et al.
Veröffentlicht: (2026)
von: Yuan, Jessie, et al.
Veröffentlicht: (2026)
AnySafe: Adapting Latent Safety Filters at Runtime via Safety Constraint Parameterization in the Latent Space
von: Agrawal, Sankalp, et al.
Veröffentlicht: (2025)
von: Agrawal, Sankalp, et al.
Veröffentlicht: (2025)
Coordinated Diffusion: Generating Multi-Agent Behavior Without Multi-Agent Demonstrations
von: Peters, Lasse, et al.
Veröffentlicht: (2026)
von: Peters, Lasse, et al.
Veröffentlicht: (2026)
Agent-to-Sim: Learning Interactive Behavior Models from Casual Longitudinal Videos
von: Yang, Gengshan, et al.
Veröffentlicht: (2024)
von: Yang, Gengshan, et al.
Veröffentlicht: (2024)
What Matters to You? Towards Visual Representation Alignment for Robot Learning
von: Tian, Ran, et al.
Veröffentlicht: (2023)
von: Tian, Ran, et al.
Veröffentlicht: (2023)
Robots that Learn to Safely Influence via Prediction-Informed Reach-Avoid Dynamic Games
von: Pandya, Ravi, et al.
Veröffentlicht: (2024)
von: Pandya, Ravi, et al.
Veröffentlicht: (2024)
What You Don't Know Can Hurt You: How Well do Latent Safety Filters Understand Partially Observable Safety Constraints?
von: Kim, Matthew, et al.
Veröffentlicht: (2025)
von: Kim, Matthew, et al.
Veröffentlicht: (2025)
Do What You Say: Steering Vision-Language-Action Models via Runtime Reasoning-Action Alignment Verification
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
Reimagination with Test-time Observation Interventions: Distractor-Robust World Model Predictions for Visual Model Predictive Control
von: Chen, Yuxin, et al.
Veröffentlicht: (2025)
von: Chen, Yuxin, et al.
Veröffentlicht: (2025)
Robots that Suggest Safe Alternatives
von: Jeong, Hyun Joe, et al.
Veröffentlicht: (2024)
von: Jeong, Hyun Joe, et al.
Veröffentlicht: (2024)
Embodied Learning of Reward for Musculoskeletal Control with Vision Language Models
von: Soedarmadji, Saraswati, et al.
Veröffentlicht: (2025)
von: Soedarmadji, Saraswati, et al.
Veröffentlicht: (2025)
Generalizing Safety Beyond Collision-Avoidance via Latent-Space Reachability Analysis
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2025)
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2025)
BadRobot: Jailbreaking Embodied LLMs in the Physical World
von: Zhang, Hangtao, et al.
Veröffentlicht: (2024)
von: Zhang, Hangtao, et al.
Veröffentlicht: (2024)
Reinforced Embodied Planning with Verifiable Reward for Real-World Robotic Manipulation
von: Bo, Zitong, et al.
Veröffentlicht: (2025)
von: Bo, Zitong, et al.
Veröffentlicht: (2025)
Uncertainty-aware Latent Safety Filters for Avoiding Out-of-Distribution Failures
von: Seo, Junwon, et al.
Veröffentlicht: (2025)
von: Seo, Junwon, et al.
Veröffentlicht: (2025)
How Good are Foundation Models in Step-by-Step Embodied Reasoning?
von: Dissanayake, Dinura, et al.
Veröffentlicht: (2025)
von: Dissanayake, Dinura, et al.
Veröffentlicht: (2025)
Conformalized Teleoperation: Confidently Mapping Human Inputs to High-Dimensional Robot Actions
von: Zhao, Michelle, et al.
Veröffentlicht: (2024)
von: Zhao, Michelle, et al.
Veröffentlicht: (2024)
StROL: Stabilized and Robust Online Learning from Humans
von: Mehta, Shaunak A., et al.
Veröffentlicht: (2023)
von: Mehta, Shaunak A., et al.
Veröffentlicht: (2023)
Updating Robot Safety Representations Online from Natural Language Feedback
von: Santos, Leonardo, et al.
Veröffentlicht: (2024)
von: Santos, Leonardo, et al.
Veröffentlicht: (2024)
Adapting by Analogy: OOD Generalization of Visuomotor Policies via Functional Correspondence
von: Gupta, Pranay, et al.
Veröffentlicht: (2025)
von: Gupta, Pranay, et al.
Veröffentlicht: (2025)
Good in Bad (GiB): Sifting Through End-user Demonstrations for Learning a Better Policy
von: Sojib, Noushad, et al.
Veröffentlicht: (2026)
von: Sojib, Noushad, et al.
Veröffentlicht: (2026)
ReWorld: Multi-Dimensional Reward Modeling for Embodied World Models
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
Embodied Navigation Foundation Model
von: Zhang, Jiazhao, et al.
Veröffentlicht: (2025)
von: Zhang, Jiazhao, et al.
Veröffentlicht: (2025)
Data Assessment for Embodied Intelligence
von: Xiao, Jiahao, et al.
Veröffentlicht: (2025)
von: Xiao, Jiahao, et al.
Veröffentlicht: (2025)
StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement
von: Seo, Junwon, et al.
Veröffentlicht: (2026)
von: Seo, Junwon, et al.
Veröffentlicht: (2026)
PathPainter: Transferring the Generalization Ability of Image Generation Models to Embodied Navigation
von: Wang, Yijin, et al.
Veröffentlicht: (2026)
von: Wang, Yijin, et al.
Veröffentlicht: (2026)
AToM-Bot: Embodied Fulfillment of Unspoken Human Needs with Affective Theory of Mind
von: Ding, Wei, et al.
Veröffentlicht: (2024)
von: Ding, Wei, et al.
Veröffentlicht: (2024)
How to Train Your Latent Control Barrier Function: Smooth Safety Filtering Under Hard-to-Model Constraints
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2025)
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2025)
Position: Olfaction Standardization is Essential for the Advancement of Embodied Artificial Intelligence
von: France, Kordel K., et al.
Veröffentlicht: (2025)
von: France, Kordel K., et al.
Veröffentlicht: (2025)
BadNAVer: Exploring Jailbreak Attacks On Vision-and-Language Navigation
von: Lyu, Wenqi, et al.
Veröffentlicht: (2025)
von: Lyu, Wenqi, et al.
Veröffentlicht: (2025)
Your Learned Constraint is Secretly a Backward Reachable Tube
von: Qadri, Mohamad, et al.
Veröffentlicht: (2025)
von: Qadri, Mohamad, et al.
Veröffentlicht: (2025)
Dual-Granularity Contrastive Reward via Generated Episodic Guidance for Efficient Embodied RL
von: Liu, Xin, et al.
Veröffentlicht: (2026)
von: Liu, Xin, et al.
Veröffentlicht: (2026)
EmbodiedCoder: Parameterized Embodied Mobile Manipulation via Modern Coding Model
von: Lin, Zefu, et al.
Veröffentlicht: (2025)
von: Lin, Zefu, et al.
Veröffentlicht: (2025)
Position: Embodied AI Requires a Privacy-Utility Trade-off
von: Fan, Xiaoliang, et al.
Veröffentlicht: (2026)
von: Fan, Xiaoliang, et al.
Veröffentlicht: (2026)
An Atomic Skill Library Construction Method for Data-Efficient Embodied Manipulation
von: Li, Dongjiang, et al.
Veröffentlicht: (2025)
von: Li, Dongjiang, et al.
Veröffentlicht: (2025)
Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning
von: Liang, Wenlong, et al.
Veröffentlicht: (2025)
von: Liang, Wenlong, et al.
Veröffentlicht: (2025)
Gain Tuning Is Not What You Need: Reward Gain Adaptation for Constrained Locomotion Learning
von: Srisuchinnawong, Arthicha, et al.
Veröffentlicht: (2025)
von: Srisuchinnawong, Arthicha, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
From Foresight to Forethought: VLM-In-the-Loop Policy Steering via Latent Alignment
von: Wu, Yilin, et al.
Veröffentlicht: (2025) -
Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment
von: Tian, Ran, et al.
Veröffentlicht: (2024) -
Not All Errors Are Made Equal: A Regret Metric for Detecting System-level Trajectory Prediction Failures
von: Nakamura, Kensuke, et al.
Veröffentlicht: (2024) -
When to Act, Ask, or Learn: Uncertainty-Aware Policy Steering
von: Yuan, Jessie, et al.
Veröffentlicht: (2026) -
AnySafe: Adapting Latent Safety Filters at Runtime via Safety Constraint Parameterization in the Latent Space
von: Agrawal, Sankalp, et al.
Veröffentlicht: (2025)