Adapting Image-based RL Policies via Predicted Rewards
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Weiyao, Fang, Xinyuan, Hager, Gregory D. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Domain Adaptation of Visual Policies with a Single Demonstration
von: Wang, Weiyao, et al.
Veröffentlicht: (2024)
von: Wang, Weiyao, et al.
Veröffentlicht: (2024)
VIHE: Virtual In-Hand Eye Transformer for 3D Robotic Manipulation
von: Wang, Weiyao, et al.
Veröffentlicht: (2024)
von: Wang, Weiyao, et al.
Veröffentlicht: (2024)
DiffusionRL: Efficient Training of Diffusion Policies for Robotic Grasping Using RL-Adapted Large-Scale Datasets
von: Makarova, Maria, et al.
Veröffentlicht: (2025)
von: Makarova, Maria, et al.
Veröffentlicht: (2025)
Adapt2Reward: Adapting Video-Language Models to Generalizable Robotic Rewards via Failure Prompts
von: Yang, Yanting, et al.
Veröffentlicht: (2024)
von: Yang, Yanting, et al.
Veröffentlicht: (2024)
CaRL: Learning Scalable Planning Policies with Simple Rewards
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2025)
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2025)
Back to Newton's Laws: Learning Vision-based Agile Flight via Differentiable Physics
von: Zhang, Yuang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuang, et al.
Veröffentlicht: (2024)
FLoRA: Sample-Efficient Preference-based RL via Low-Rank Style Adaptation of Reward Functions
von: Marta, Daniel, et al.
Veröffentlicht: (2025)
von: Marta, Daniel, et al.
Veröffentlicht: (2025)
COMPASS: Cross-embodiment Mobility Policy via Residual RL and Skill Synthesis
von: Liu, Wei, et al.
Veröffentlicht: (2025)
von: Liu, Wei, et al.
Veröffentlicht: (2025)
Dual-Granularity Contrastive Reward via Generated Episodic Guidance for Efficient Embodied RL
von: Liu, Xin, et al.
Veröffentlicht: (2026)
von: Liu, Xin, et al.
Veröffentlicht: (2026)
Q-Guided Stein Variational Model Predictive Control via RL-informed Policy Prior
von: Cai, Shizhe, et al.
Veröffentlicht: (2025)
von: Cai, Shizhe, et al.
Veröffentlicht: (2025)
GRAPPA: Generalizing and Adapting Robot Policies via Online Agentic Guidance
von: Bucker, Arthur, et al.
Veröffentlicht: (2024)
von: Bucker, Arthur, et al.
Veröffentlicht: (2024)
A Heuristic Approach for Performance Tuning in RL-based Quadrotor Control via Reward Design and Termination Conditions
von: Suarez, Fausto Mauricio Lagos, et al.
Veröffentlicht: (2026)
von: Suarez, Fausto Mauricio Lagos, et al.
Veröffentlicht: (2026)
Driving Beyond Privilege: Distilling Dense-Reward Knowledge into Sparse-Reward Policies
von: Khanzada, Feeza Khan, et al.
Veröffentlicht: (2025)
von: Khanzada, Feeza Khan, et al.
Veröffentlicht: (2025)
Integrating Model-based Control and RL for Sim2Real Transfer of Tight Insertion Policies
von: Marougkas, Isidoros, et al.
Veröffentlicht: (2025)
von: Marougkas, Isidoros, et al.
Veröffentlicht: (2025)
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
von: Zhang, Chen Bo Calvin, et al.
Veröffentlicht: (2024)
von: Zhang, Chen Bo Calvin, et al.
Veröffentlicht: (2024)
Rapidly Adapting Policies to the Real World via Simulation-Guided Fine-Tuning
von: Yin, Patrick, et al.
Veröffentlicht: (2025)
von: Yin, Patrick, et al.
Veröffentlicht: (2025)
Recovering Hidden Reward in Diffusion-Based Policies
von: Ji, Yanbiao, et al.
Veröffentlicht: (2026)
von: Ji, Yanbiao, et al.
Veröffentlicht: (2026)
Residual Off-Policy RL for Finetuning Behavior Cloning Policies
von: Ankile, Lars, et al.
Veröffentlicht: (2025)
von: Ankile, Lars, et al.
Veröffentlicht: (2025)
Robust RL with LLM-Driven Data Synthesis and Policy Adaptation for Autonomous Driving
von: Wu, Sihao, et al.
Veröffentlicht: (2024)
von: Wu, Sihao, et al.
Veröffentlicht: (2024)
Enhancing Exploration with Diffusion Policies in Hybrid Off-Policy RL: Application to Non-Prehensile Manipulation
von: Le, Huy, et al.
Veröffentlicht: (2024)
von: Le, Huy, et al.
Veröffentlicht: (2024)
Adapting Neural Robot Dynamics on the Fly for Predictive Control
von: Altawaitan, Abdullah, et al.
Veröffentlicht: (2026)
von: Altawaitan, Abdullah, et al.
Veröffentlicht: (2026)
Data-Efficient Policy Selection for Navigation in Partial Maps via Subgoal-Based Abstraction
von: Paudel, Abhishek, et al.
Veröffentlicht: (2023)
von: Paudel, Abhishek, et al.
Veröffentlicht: (2023)
History-Aware Visuomotor Policy Learning via Point Tracking
von: Chen, Jingjing, et al.
Veröffentlicht: (2025)
von: Chen, Jingjing, et al.
Veröffentlicht: (2025)
VLA-OPD: Bridging Offline SFT and Online RL for Vision-Language-Action Models via On-Policy Distillation
von: Zhong, Zhide, et al.
Veröffentlicht: (2026)
von: Zhong, Zhide, et al.
Veröffentlicht: (2026)
LocoVLM: Grounding Vision and Language for Adapting Versatile Legged Locomotion Policies
von: Nahrendra, I Made Aswin, et al.
Veröffentlicht: (2026)
von: Nahrendra, I Made Aswin, et al.
Veröffentlicht: (2026)
Efficient Language-instructed Skill Acquisition via Reward-Policy Co-Evolution
von: Huang, Changxin, et al.
Veröffentlicht: (2024)
von: Huang, Changxin, et al.
Veröffentlicht: (2024)
TopoCut: Learning Multi-Step Cutting with Spectral Rewards and Discrete Diffusion Policies
von: Wang, Liquan, et al.
Veröffentlicht: (2025)
von: Wang, Liquan, et al.
Veröffentlicht: (2025)
A comparison of RL-based and PID controllers for 6-DOF swimming robots: hybrid underwater object tracking
von: Lotfi, Faraz, et al.
Veröffentlicht: (2024)
von: Lotfi, Faraz, et al.
Veröffentlicht: (2024)
RoboScape-R: Unified Reward-Observation World Models for Generalizable Robotics Training via RL
von: Tang, Yinzhou, et al.
Veröffentlicht: (2025)
von: Tang, Yinzhou, et al.
Veröffentlicht: (2025)
End-to-end RL Improves Dexterous Grasping Policies
von: Singh, Ritvik, et al.
Veröffentlicht: (2025)
von: Singh, Ritvik, et al.
Veröffentlicht: (2025)
Inference-stage Adaptation-projection Strategy Adapts Diffusion Policy to Cross-manipulators Scenarios
von: Yao, Xiangtong, et al.
Veröffentlicht: (2025)
von: Yao, Xiangtong, et al.
Veröffentlicht: (2025)
Rewarding DINO: Predicting Dense Rewards with Vision Foundation Models
von: Krack, Pierre, et al.
Veröffentlicht: (2026)
von: Krack, Pierre, et al.
Veröffentlicht: (2026)
Safe Policy Exploration Improvement via Subgoals
von: Angulo, Brian, et al.
Veröffentlicht: (2024)
von: Angulo, Brian, et al.
Veröffentlicht: (2024)
120 Minutes and a Laptop: Minimalist Image-goal Navigation via Unsupervised Exploration and Offline RL
von: Liu, Xiaoming, et al.
Veröffentlicht: (2026)
von: Liu, Xiaoming, et al.
Veröffentlicht: (2026)
Adapting by Analogy: OOD Generalization of Visuomotor Policies via Functional Correspondence
von: Gupta, Pranay, et al.
Veröffentlicht: (2025)
von: Gupta, Pranay, et al.
Veröffentlicht: (2025)
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction
von: Fan, Chenyou, et al.
Veröffentlicht: (2025)
von: Fan, Chenyou, et al.
Veröffentlicht: (2025)
DyPNIPP: Predicting Environment Dynamics for RL-based Robust Informative Path Planning
von: Deolasee, Srujan, et al.
Veröffentlicht: (2024)
von: Deolasee, Srujan, et al.
Veröffentlicht: (2024)
Refined Policy Distillation: From VLA Generalists to RL Experts
von: Jülg, Tobias, et al.
Veröffentlicht: (2025)
von: Jülg, Tobias, et al.
Veröffentlicht: (2025)
Leveraging Symmetry in RL-based Legged Locomotion Control
von: Su, Zhi, et al.
Veröffentlicht: (2024)
von: Su, Zhi, et al.
Veröffentlicht: (2024)
POLICEd RL: Learning Closed-Loop Robot Control Policies with Provable Satisfaction of Hard Constraints
von: Bouvier, Jean-Baptiste, et al.
Veröffentlicht: (2024)
von: Bouvier, Jean-Baptiste, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Domain Adaptation of Visual Policies with a Single Demonstration
von: Wang, Weiyao, et al.
Veröffentlicht: (2024) -
VIHE: Virtual In-Hand Eye Transformer for 3D Robotic Manipulation
von: Wang, Weiyao, et al.
Veröffentlicht: (2024) -
DiffusionRL: Efficient Training of Diffusion Policies for Robotic Grasping Using RL-Adapted Large-Scale Datasets
von: Makarova, Maria, et al.
Veröffentlicht: (2025) -
Adapt2Reward: Adapting Video-Language Models to Generalizable Robotic Rewards via Failure Prompts
von: Yang, Yanting, et al.
Veröffentlicht: (2024) -
CaRL: Learning Scalable Planning Policies with Simple Rewards
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2025)