LLM-Personalize: Aligning LLM Planners with Human Preferences via Reinforced Self-Training for Housekeeping Robots
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Han, Dongge, McInroe, Trevor, Jelley, Adam, Albrecht, Stefano V., Bell, Peter, Storkey, Amos |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Planning to Go Out-of-Distribution in Offline-to-Online Reinforcement Learning
von: McInroe, Trevor, et al.
Veröffentlicht: (2023)
von: McInroe, Trevor, et al.
Veröffentlicht: (2023)
Efficient Offline Reinforcement Learning: First Imitate, then Improve
von: Jelley, Adam, et al.
Veröffentlicht: (2024)
von: Jelley, Adam, et al.
Veröffentlicht: (2024)
Object-Centric World Models from Few-Shot Annotations for Sample-Efficient Reinforcement Learning
von: Zhang, Weipu, et al.
Veröffentlicht: (2025)
von: Zhang, Weipu, et al.
Veröffentlicht: (2025)
Enhancing Tactile-based Reinforcement Learning for Robotic Control
von: Miller, Elle, et al.
Veröffentlicht: (2025)
von: Miller, Elle, et al.
Veröffentlicht: (2025)
Multi-Horizon Representations with Hierarchical Forward Models for Reinforcement Learning
von: McInroe, Trevor, et al.
Veröffentlicht: (2022)
von: McInroe, Trevor, et al.
Veröffentlicht: (2022)
Terra Nova: A Comprehensive Challenge Environment for Intelligent Agents
von: McInroe, Trevor
Veröffentlicht: (2025)
von: McInroe, Trevor
Veröffentlicht: (2025)
Assistax: A Hardware-Accelerated Reinforcement Learning Benchmark for Assistive Robotics
von: Hinckeldey, Leonard, et al.
Veröffentlicht: (2025)
von: Hinckeldey, Leonard, et al.
Veröffentlicht: (2025)
roto 2.0: The Robot Tactile Olympiad
von: Miller, Elle, et al.
Veröffentlicht: (2026)
von: Miller, Elle, et al.
Veröffentlicht: (2026)
PixelBrax: Learning Continuous Control from Pixels End-to-End on the GPU
von: McInroe, Trevor, et al.
Veröffentlicht: (2025)
von: McInroe, Trevor, et al.
Veröffentlicht: (2025)
Forgetting is Everywhere
von: Sanati, Ben, et al.
Veröffentlicht: (2025)
von: Sanati, Ben, et al.
Veröffentlicht: (2025)
HiCRISP: An LLM-based Hierarchical Closed-Loop Robotic Intelligent Self-Correction Planner
von: Ming, Chenlin, et al.
Veröffentlicht: (2023)
von: Ming, Chenlin, et al.
Veröffentlicht: (2023)
Preference Aligned Diffusion Planner for Quadrupedal Locomotion Control
von: Yuan, Xinyi, et al.
Veröffentlicht: (2024)
von: Yuan, Xinyi, et al.
Veröffentlicht: (2024)
Aligning Agents like Large Language Models
von: Jelley, Adam, et al.
Veröffentlicht: (2024)
von: Jelley, Adam, et al.
Veröffentlicht: (2024)
T3 Planner: A Self-Correcting LLM Framework for Robotic Motion Planning with Temporal Logic
von: Li, Jia, et al.
Veröffentlicht: (2025)
von: Li, Jia, et al.
Veröffentlicht: (2025)
Collision- and Reachability-Aware Multi-Robot Control with Grounded LLM Planners
von: Ji, Jiabao, et al.
Veröffentlicht: (2025)
von: Ji, Jiabao, et al.
Veröffentlicht: (2025)
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation
von: Seneviratne, Gershom, et al.
Veröffentlicht: (2025)
von: Seneviratne, Gershom, et al.
Veröffentlicht: (2025)
InteLiPlan: An Interactive Lightweight LLM-Based Planner for Domestic Robot Autonomy
von: Ly, Kim Tien, et al.
Veröffentlicht: (2024)
von: Ly, Kim Tien, et al.
Veröffentlicht: (2024)
LLM-as-BT-Planner: Leveraging LLMs for Behavior Tree Generation in Robot Task Planning
von: Ao, Jicong, et al.
Veröffentlicht: (2024)
von: Ao, Jicong, et al.
Veröffentlicht: (2024)
Decentralized Intent-Based Multi-Robot Task Planner with LLM Oracles on Hyperledger Fabric
von: Keramat, Farhad, et al.
Veröffentlicht: (2026)
von: Keramat, Farhad, et al.
Veröffentlicht: (2026)
Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning
von: Garcin, Samuel, et al.
Veröffentlicht: (2025)
von: Garcin, Samuel, et al.
Veröffentlicht: (2025)
A Task-Efficient Reinforcement Learning Task-Motion Planner for Safe Human-Robot Cooperation
von: Liu, Gaoyuan, et al.
Veröffentlicht: (2025)
von: Liu, Gaoyuan, et al.
Veröffentlicht: (2025)
Aligning Robot Navigation Behaviors with Human Intentions and Preferences
von: Karnan, Haresh
Veröffentlicht: (2024)
von: Karnan, Haresh
Veröffentlicht: (2024)
CorrectionPlanner: Self-Correction Planner with Reinforcement Learning in Autonomous Driving
von: Guo, Yihong, et al.
Veröffentlicht: (2026)
von: Guo, Yihong, et al.
Veröffentlicht: (2026)
Energy-Efficient Motion Planner for Legged Robots
von: Schperberg, Alexander, et al.
Veröffentlicht: (2025)
von: Schperberg, Alexander, et al.
Veröffentlicht: (2025)
Personalization in Human-Robot Interaction through Preference-based Action Representation Learning
von: Wang, Ruiqi, et al.
Veröffentlicht: (2024)
von: Wang, Ruiqi, et al.
Veröffentlicht: (2024)
Latent Embedding Adaptation for Human Preference Alignment in Diffusion Planners
von: Ng, Wen Zheng Terence, et al.
Veröffentlicht: (2025)
von: Ng, Wen Zheng Terence, et al.
Veröffentlicht: (2025)
Aligning Humans and Robots via Reinforcement Learning from Implicit Human Feedback
von: Kim, Suzie, et al.
Veröffentlicht: (2025)
von: Kim, Suzie, et al.
Veröffentlicht: (2025)
Fast Trajectory Planner with a Reinforcement Learning-based Controller for Robotic Manipulators
von: Wang, Yongliang, et al.
Veröffentlicht: (2025)
von: Wang, Yongliang, et al.
Veröffentlicht: (2025)
Robo-Troj: Attacking LLM-based Task Planners
von: Nahian, Mohaiminul Al, et al.
Veröffentlicht: (2025)
von: Nahian, Mohaiminul Al, et al.
Veröffentlicht: (2025)
OHP-RL: Online Human Preference as Guidance in Reinforcement Learning for Robot Manipulation
von: Mo, Yunyang, et al.
Veröffentlicht: (2026)
von: Mo, Yunyang, et al.
Veröffentlicht: (2026)
LBAP: Improved Uncertainty Alignment of LLM Planners using Bayesian Inference
von: Mullen Jr., James F., et al.
Veröffentlicht: (2024)
von: Mullen Jr., James F., et al.
Veröffentlicht: (2024)
KnowDiffuser: A Knowledge-Guided Diffusion Planner with LLM Reasoning
von: Ding, Fan, et al.
Veröffentlicht: (2026)
von: Ding, Fan, et al.
Veröffentlicht: (2026)
MimicDreamer: Aligning Human and Robot Demonstrations for Scalable VLA Training
von: Li, Haoyun, et al.
Veröffentlicht: (2025)
von: Li, Haoyun, et al.
Veröffentlicht: (2025)
Learning Human-Robot Handshaking Preferences for Quadruped Robots
von: Chappuis, Alessandra, et al.
Veröffentlicht: (2024)
von: Chappuis, Alessandra, et al.
Veröffentlicht: (2024)
Reinforcement Learning from Implicit Neural Feedback for Human-Aligned Robot Control
von: Kim, Suzie
Veröffentlicht: (2025)
von: Kim, Suzie
Veröffentlicht: (2025)
Human Leading or Following Preferences: Effects on Human Perception of the Robot and the Human-Robot Collaboration
von: Noormohammadi-Asl, Ali, et al.
Veröffentlicht: (2024)
von: Noormohammadi-Asl, Ali, et al.
Veröffentlicht: (2024)
Efficient Reinforcement Learning of Task Planners for Robotic Palletization through Iterative Action Masking Learning
von: Wu, Zheng, et al.
Veröffentlicht: (2024)
von: Wu, Zheng, et al.
Veröffentlicht: (2024)
Safe Planner: Empowering Safety Awareness in Large Pre-Trained Models for Robot Task Planning
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
Evaluating Uncertainty-based Failure Detection for Closed-Loop LLM Planners
von: Zheng, Zhi, et al.
Veröffentlicht: (2024)
von: Zheng, Zhi, et al.
Veröffentlicht: (2024)
LLM-based Human-like Traffic Simulation for Self-driving Tests
von: Li, Wendi, et al.
Veröffentlicht: (2025)
von: Li, Wendi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Planning to Go Out-of-Distribution in Offline-to-Online Reinforcement Learning
von: McInroe, Trevor, et al.
Veröffentlicht: (2023) -
Efficient Offline Reinforcement Learning: First Imitate, then Improve
von: Jelley, Adam, et al.
Veröffentlicht: (2024) -
Object-Centric World Models from Few-Shot Annotations for Sample-Efficient Reinforcement Learning
von: Zhang, Weipu, et al.
Veröffentlicht: (2025) -
Enhancing Tactile-based Reinforcement Learning for Robotic Control
von: Miller, Elle, et al.
Veröffentlicht: (2025) -
Multi-Horizon Representations with Hierarchical Forward Models for Reinforcement Learning
von: McInroe, Trevor, et al.
Veröffentlicht: (2022)