RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Yufei, Sun, Zhanyi, Zhang, Jesse, Xian, Zhou, Biyik, Erdem, Held, David, Erickson, Zackory |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Real-World Offline Reinforcement Learning from Vision Language Model Feedback
di: Venkataraman, Sreyas, et al.
Pubblicazione: (2024)
di: Venkataraman, Sreyas, et al.
Pubblicazione: (2024)
DiffTORI: Differentiable Trajectory Optimization for Deep Reinforcement and Imitation Learning
di: Wan, Weikang, et al.
Pubblicazione: (2024)
di: Wan, Weikang, et al.
Pubblicazione: (2024)
Force-Constrained Visual Policy: Safe Robot-Assisted Dressing via Multi-Modal Sensing
di: Sun, Zhanyi, et al.
Pubblicazione: (2023)
di: Sun, Zhanyi, et al.
Pubblicazione: (2023)
MILE: Model-based Intervention Learning
di: Korkmaz, Yigit, et al.
Pubblicazione: (2025)
di: Korkmaz, Yigit, et al.
Pubblicazione: (2025)
Causally Robust Reward Learning from Reason-Augmented Preference Feedback
di: Hwang, Minjune, et al.
Pubblicazione: (2026)
di: Hwang, Minjune, et al.
Pubblicazione: (2026)
Force-Modulated Visual Policy for Robot-Assisted Dressing with Arm Motions
di: Hao, Alexis Yihong, et al.
Pubblicazione: (2025)
di: Hao, Alexis Yihong, et al.
Pubblicazione: (2025)
Geometric Red-Teaming for Robotic Manipulation
di: Goel, Divyam, et al.
Pubblicazione: (2025)
di: Goel, Divyam, et al.
Pubblicazione: (2025)
ViSaRL: Visual Reinforcement Learning Guided by Human Saliency
di: Liang, Anthony, et al.
Pubblicazione: (2024)
di: Liang, Anthony, et al.
Pubblicazione: (2024)
IMPACT: Intelligent Motion Planning with Acceptable Contact Trajectories via Vision-Language Models
di: Ling, Yiyang, et al.
Pubblicazione: (2025)
di: Ling, Yiyang, et al.
Pubblicazione: (2025)
ArticuBot: Learning Universal Articulated Object Manipulation Policy via Large Scale Simulation
di: Wang, Yufei, et al.
Pubblicazione: (2025)
di: Wang, Yufei, et al.
Pubblicazione: (2025)
RoboGen: Towards Unleashing Infinite Data for Automated Robot Learning via Generative Simulation
di: Wang, Yufei, et al.
Pubblicazione: (2023)
di: Wang, Yufei, et al.
Pubblicazione: (2023)
Batch Active Learning of Reward Functions from Human Preferences
di: Bıyık, Erdem, et al.
Pubblicazione: (2024)
di: Bıyık, Erdem, et al.
Pubblicazione: (2024)
GABRIL: Gaze-Based Regularization for Mitigating Causal Confusion in Imitation Learning
di: Banayeeanzade, Amin, et al.
Pubblicazione: (2025)
di: Banayeeanzade, Amin, et al.
Pubblicazione: (2025)
VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving
di: Huang, Zilin, et al.
Pubblicazione: (2024)
di: Huang, Zilin, et al.
Pubblicazione: (2024)
DriveVLM-RL: Neuroscience-Inspired Reinforcement Learning with Vision-Language Models for Safe and Deployable Autonomous Driving
di: Huang, Zilin, et al.
Pubblicazione: (2026)
di: Huang, Zilin, et al.
Pubblicazione: (2026)
CLAM: Continuous Latent Action Models for Robot Learning from Unlabeled Demonstrations
di: Liang, Anthony, et al.
Pubblicazione: (2025)
di: Liang, Anthony, et al.
Pubblicazione: (2025)
EDO-Net: Learning Elastic Properties of Deformable Objects from Graph Dynamics
di: Longhini, Alberta, et al.
Pubblicazione: (2022)
di: Longhini, Alberta, et al.
Pubblicazione: (2022)
VLM-SAFE: Vision-Language Model-Guided Safety-Aware Reinforcement Learning with World Models for Autonomous Driving
di: Qu, Yansong, et al.
Pubblicazione: (2025)
di: Qu, Yansong, et al.
Pubblicazione: (2025)
EXTRACT: Efficient Policy Learning by Extracting Transferable Robot Skills from Offline Data
di: Zhang, Jesse, et al.
Pubblicazione: (2024)
di: Zhang, Jesse, et al.
Pubblicazione: (2024)
Actor-Free Continuous Control via Structurally Maximizable Q-Functions
di: Korkmaz, Yigit, et al.
Pubblicazione: (2025)
di: Korkmaz, Yigit, et al.
Pubblicazione: (2025)
When a Robot is More Capable than a Human: Learning from Constrained Demonstrators
di: Li, Xinhu, et al.
Pubblicazione: (2025)
di: Li, Xinhu, et al.
Pubblicazione: (2025)
Discover, Learn, and Reinforce: Scaling Vision-Language-Action Pretraining with Diverse RL-Generated Trajectories
di: Yang, Rushuai, et al.
Pubblicazione: (2025)
di: Yang, Rushuai, et al.
Pubblicazione: (2025)
RAILGUN: A Unified Convolutional Policy for Multi-Agent Path Finding Across Different Environments and Tasks
di: Tang, Yimin, et al.
Pubblicazione: (2025)
di: Tang, Yimin, et al.
Pubblicazione: (2025)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
di: Zhang, Jesse, et al.
Pubblicazione: (2025)
di: Zhang, Jesse, et al.
Pubblicazione: (2025)
A Generalized Acquisition Function for Preference-based Reward Learning
di: Ellis, Evan, et al.
Pubblicazione: (2024)
di: Ellis, Evan, et al.
Pubblicazione: (2024)
Reinforcement Learning in a Safety-Embedded MDP with Trajectory Optimization
di: Yang, Fan, et al.
Pubblicazione: (2023)
di: Yang, Fan, et al.
Pubblicazione: (2023)
On Time-Indexing as Inductive Bias in Deep RL for Sequential Manipulation Tasks
di: Qureshi, M. Nomaan, et al.
Pubblicazione: (2024)
di: Qureshi, M. Nomaan, et al.
Pubblicazione: (2024)
Multi-Agent Inverse Q-Learning from Demonstrations
di: Haynam, Nathaniel, et al.
Pubblicazione: (2025)
di: Haynam, Nathaniel, et al.
Pubblicazione: (2025)
LAMS: LLM-Driven Automatic Mode Switching for Assistive Teleoperation
di: Tao, Yiran, et al.
Pubblicazione: (2025)
di: Tao, Yiran, et al.
Pubblicazione: (2025)
VLM-RRT: Vision Language Model Guided RRT Search for Autonomous UAV Navigation
di: Ye, Jianlin, et al.
Pubblicazione: (2025)
di: Ye, Jianlin, et al.
Pubblicazione: (2025)
VLM-DEWM: Dynamic External World Model for Verifiable and Resilient Vision-Language Planning in Manufacturing
di: Tang, Guoqin, et al.
Pubblicazione: (2026)
di: Tang, Guoqin, et al.
Pubblicazione: (2026)
COMRES-VLM: Coordinated Multi-Robot Exploration and Search using Vision Language Models
di: Wang, Ruiyang, et al.
Pubblicazione: (2025)
di: Wang, Ruiyang, et al.
Pubblicazione: (2025)
A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning
di: Zhai, Shaopeng, et al.
Pubblicazione: (2025)
di: Zhai, Shaopeng, et al.
Pubblicazione: (2025)
AppleVLM: End-to-end Autonomous Driving with Advanced Perception and Planning-Enhanced Vision-Language Models
di: Han, Yuxuan, et al.
Pubblicazione: (2026)
di: Han, Yuxuan, et al.
Pubblicazione: (2026)
World4RL: Diffusion World Models for Policy Refinement with Reinforcement Learning for Robotic Manipulation
di: Jiang, Zhennan, et al.
Pubblicazione: (2025)
di: Jiang, Zhennan, et al.
Pubblicazione: (2025)
D-VLA: A High-Concurrency Distributed Asynchronous Reinforcement Learning Framework for Vision-Language-Action Models
di: Guo, Yucheng, et al.
Pubblicazione: (2026)
di: Guo, Yucheng, et al.
Pubblicazione: (2026)
RLRC: Reinforcement Learning-based Recovery for Compressed Vision-Language-Action Models
di: Chen, Yuxuan, et al.
Pubblicazione: (2025)
di: Chen, Yuxuan, et al.
Pubblicazione: (2025)
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
di: Lu, Guanxing, et al.
Pubblicazione: (2025)
di: Lu, Guanxing, et al.
Pubblicazione: (2025)
EXPO-FT: Sample-Efficient Reinforcement Learning Finetuning for Vision-Language-Action Models
di: Dong, Perry, et al.
Pubblicazione: (2026)
di: Dong, Perry, et al.
Pubblicazione: (2026)
MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models
di: Zhao, Han, et al.
Pubblicazione: (2025)
di: Zhao, Han, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Real-World Offline Reinforcement Learning from Vision Language Model Feedback
di: Venkataraman, Sreyas, et al.
Pubblicazione: (2024) -
DiffTORI: Differentiable Trajectory Optimization for Deep Reinforcement and Imitation Learning
di: Wan, Weikang, et al.
Pubblicazione: (2024) -
Force-Constrained Visual Policy: Safe Robot-Assisted Dressing via Multi-Modal Sensing
di: Sun, Zhanyi, et al.
Pubblicazione: (2023) -
MILE: Model-based Intervention Learning
di: Korkmaz, Yigit, et al.
Pubblicazione: (2025) -
Causally Robust Reward Learning from Reason-Augmented Preference Feedback
di: Hwang, Minjune, et al.
Pubblicazione: (2026)