PrefCLM: Enhancing Preference-based Reinforcement Learning with Crowdsourced Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Ruiqi, Zhao, Dezhong, Yuan, Ziqin, Obi, Ike, Min, Byung-Cheol |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PrefMoE: Robust Preference Modeling with Mixture-of-Experts Reward Learning
von: Yuan, Ziqin, et al.
Veröffentlicht: (2026)
von: Yuan, Ziqin, et al.
Veröffentlicht: (2026)
PrefMMT: Modeling Human Preferences in Preference-based Reinforcement Learning with Multimodal Transformers
von: Zhao, Dezhong, et al.
Veröffentlicht: (2024)
von: Zhao, Dezhong, et al.
Veröffentlicht: (2024)
Adaptive Task Allocation in Multi-Human Multi-Robot Teams under Team Heterogeneity and Dynamic Information Uncertainty
von: Yuan, Ziqin, et al.
Veröffentlicht: (2024)
von: Yuan, Ziqin, et al.
Veröffentlicht: (2024)
Personalization in Human-Robot Interaction through Preference-based Action Representation Learning
von: Wang, Ruiqi, et al.
Veröffentlicht: (2024)
von: Wang, Ruiqi, et al.
Veröffentlicht: (2024)
Unifying Large Language Model and Deep Reinforcement Learning for Human-in-Loop Interactive Socially-aware Navigation
von: Wang, Weizheng, et al.
Veröffentlicht: (2024)
von: Wang, Weizheng, et al.
Veröffentlicht: (2024)
PRIMT: Preference-based Reinforcement Learning with Multimodal Feedback and Trajectory Synthesis from Foundation Models
von: Wang, Ruiqi, et al.
Veröffentlicht: (2025)
von: Wang, Ruiqi, et al.
Veröffentlicht: (2025)
Modeling and Evaluating Trust Dynamics in Multi-Human Multi-Robot Task Allocation
von: Obi, Ike, et al.
Veröffentlicht: (2024)
von: Obi, Ike, et al.
Veröffentlicht: (2024)
Multi-Agent LLM Actor-Critic Framework for Social Robot Navigation
von: Wang, Weizheng, et al.
Veröffentlicht: (2025)
von: Wang, Weizheng, et al.
Veröffentlicht: (2025)
REBEL: Rule-based and Experience-enhanced Learning with LLMs for Initial Task Allocation in Multi-Human Multi-Robot Teaming
von: Gupte, Arjun, et al.
Veröffentlicht: (2024)
von: Gupte, Arjun, et al.
Veröffentlicht: (2024)
Logic-Guided Socially-aware Robot Navigation World Model
von: Wang, Weizheng, et al.
Veröffentlicht: (2025)
von: Wang, Weizheng, et al.
Veröffentlicht: (2025)
SafePlan: Leveraging Formal Logic and Chain-of-Thought Reasoning for Enhanced Safety in LLM-based Robotic Task Planning
von: Obi, Ike, et al.
Veröffentlicht: (2025)
von: Obi, Ike, et al.
Veröffentlicht: (2025)
Multi-Robot Cooperative Socially-Aware Navigation Using Multi-Agent Reinforcement Learning
von: Wang, Weizheng, et al.
Veröffentlicht: (2023)
von: Wang, Weizheng, et al.
Veröffentlicht: (2023)
Pre-Execution Safety Gate & Task Safety Contracts for LLM-Controlled Robot Systems
von: Obi, Ike, et al.
Veröffentlicht: (2026)
von: Obi, Ike, et al.
Veröffentlicht: (2026)
ZeroCAP: Zero-Shot Multi-Robot Context Aware Pattern Formation via Large Language Models
von: Venkatesh, Vishnunandan L. N., et al.
Veröffentlicht: (2024)
von: Venkatesh, Vishnunandan L. N., et al.
Veröffentlicht: (2024)
SMART-LLM: Smart Multi-Agent Robot Task Planning using Large Language Models
von: Kannan, Shyam Sundar, et al.
Veröffentlicht: (2023)
von: Kannan, Shyam Sundar, et al.
Veröffentlicht: (2023)
Hypergraph-based Coordinated Task Allocation and Socially-aware Navigation for Multi-Robot Systems
von: Wang, Weizheng, et al.
Veröffentlicht: (2024)
von: Wang, Weizheng, et al.
Veröffentlicht: (2024)
Cognitive Load-based Affective Workload Allocation for Multi-human Multi-robot Teams
von: Jo, Wonse, et al.
Veröffentlicht: (2023)
von: Jo, Wonse, et al.
Veröffentlicht: (2023)
Learning from Demonstration Framework for Multi-Robot Systems Using Interaction Keypoints and Soft Actor-Critic Methods
von: Venkatesh, Vishnunandan L. N., et al.
Veröffentlicht: (2024)
von: Venkatesh, Vishnunandan L. N., et al.
Veröffentlicht: (2024)
PlaceFormer: Transformer-based Visual Place Recognition using Multi-Scale Patch Selection and Fusion
von: Kannan, Shyam Sundar, et al.
Veröffentlicht: (2024)
von: Kannan, Shyam Sundar, et al.
Veröffentlicht: (2024)
Asynchronous Large Language Model Enhanced Planner for Autonomous Driving
von: Chen, Yuan, et al.
Veröffentlicht: (2024)
von: Chen, Yuan, et al.
Veröffentlicht: (2024)
Semantic Layering in Room Segmentation via LLMs
von: Kim, Taehyeon, et al.
Veröffentlicht: (2024)
von: Kim, Taehyeon, et al.
Veröffentlicht: (2024)
Human-Robot Cooperative Distribution Coupling for Hamiltonian-Constrained Social Navigation
von: Wang, Weizheng, et al.
Veröffentlicht: (2024)
von: Wang, Weizheng, et al.
Veröffentlicht: (2024)
REFLEX: Metacognitive Reasoning for Reflective Zero-Shot Robotic Planning with Large Language Models
von: Lin, Wenjie, et al.
Veröffentlicht: (2025)
von: Lin, Wenjie, et al.
Veröffentlicht: (2025)
ZeroSCD: Zero-Shot Street Scene Change Detection
von: Kannan, Shyam Sundar, et al.
Veröffentlicht: (2024)
von: Kannan, Shyam Sundar, et al.
Veröffentlicht: (2024)
Few-Shot Demonstration-Driven Task Coordination and Trajectory Execution for Multi-Robot Systems
von: Kim, Taehyeon, et al.
Veröffentlicht: (2025)
von: Kim, Taehyeon, et al.
Veröffentlicht: (2025)
LAPP: Large Language Model Feedback for Preference-Driven Reinforcement Learning
von: Jian, Pingcheng, et al.
Veröffentlicht: (2025)
von: Jian, Pingcheng, et al.
Veröffentlicht: (2025)
Generalizable Skill Learning for Construction Robots with Crowdsourced Natural Language Instructions, Composable Skills Standardization, and Large Language Model
von: Yu, Hongrui, et al.
Veröffentlicht: (2025)
von: Yu, Hongrui, et al.
Veröffentlicht: (2025)
Research on Robot Path Planning Based on Reinforcement Learning
von: Ruiqi, Wang
Veröffentlicht: (2024)
von: Ruiqi, Wang
Veröffentlicht: (2024)
Residual Reward Models for Preference-based Reinforcement Learning
von: Cao, Chenyang, et al.
Veröffentlicht: (2025)
von: Cao, Chenyang, et al.
Veröffentlicht: (2025)
RIME: Robust Preference-based Reinforcement Learning with Noisy Preferences
von: Cheng, Jie, et al.
Veröffentlicht: (2024)
von: Cheng, Jie, et al.
Veröffentlicht: (2024)
Enhancing Rating-Based Reinforcement Learning to Effectively Leverage Feedback from Large Vision-Language Models
von: Luu, Tung Minh, et al.
Veröffentlicht: (2025)
von: Luu, Tung Minh, et al.
Veröffentlicht: (2025)
Deep Reinforcement Learning-based Large-scale Robot Exploration
von: Cao, Yuhong, et al.
Veröffentlicht: (2024)
von: Cao, Yuhong, et al.
Veröffentlicht: (2024)
Survey on Large Language Model-Enhanced Reinforcement Learning: Concept, Taxonomy, and Methods
von: Cao, Yuji, et al.
Veröffentlicht: (2024)
von: Cao, Yuji, et al.
Veröffentlicht: (2024)
Rapidly Learning Soft Robot Control via Implicit Time-Stepping
von: Choi, Andrew, et al.
Veröffentlicht: (2025)
von: Choi, Andrew, et al.
Veröffentlicht: (2025)
Multimodal Audio-based Disease Prediction with Transformer-based Hierarchical Fusion Network
von: Cai, Jinjin, et al.
Veröffentlicht: (2024)
von: Cai, Jinjin, et al.
Veröffentlicht: (2024)
PREDILECT: Preferences Delineated with Zero-Shot Language-based Reasoning in Reinforcement Learning
von: Holk, Simon, et al.
Veröffentlicht: (2024)
von: Holk, Simon, et al.
Veröffentlicht: (2024)
PRISM: Preference Refinement via Implicit Scene Modeling for 3D Vision-Language Preference-Based Reinforcement Learning
von: Sun, Yirong, et al.
Veröffentlicht: (2025)
von: Sun, Yirong, et al.
Veröffentlicht: (2025)
RoboCrowd: Scaling Robot Data Collection through Crowdsourcing
von: Mirchandani, Suvir, et al.
Veröffentlicht: (2024)
von: Mirchandani, Suvir, et al.
Veröffentlicht: (2024)
Large Language Model guided Deep Reinforcement Learning for Decision Making in Autonomous Driving
von: Pang, Hao, et al.
Veröffentlicht: (2024)
von: Pang, Hao, et al.
Veröffentlicht: (2024)
Vision-Language Navigation with Continual Learning
von: Li, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Li, Zhiyuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
PrefMoE: Robust Preference Modeling with Mixture-of-Experts Reward Learning
von: Yuan, Ziqin, et al.
Veröffentlicht: (2026) -
PrefMMT: Modeling Human Preferences in Preference-based Reinforcement Learning with Multimodal Transformers
von: Zhao, Dezhong, et al.
Veröffentlicht: (2024) -
Adaptive Task Allocation in Multi-Human Multi-Robot Teams under Team Heterogeneity and Dynamic Information Uncertainty
von: Yuan, Ziqin, et al.
Veröffentlicht: (2024) -
Personalization in Human-Robot Interaction through Preference-based Action Representation Learning
von: Wang, Ruiqi, et al.
Veröffentlicht: (2024) -
Unifying Large Language Model and Deep Reinforcement Learning for Human-in-Loop Interactive Socially-aware Navigation
von: Wang, Weizheng, et al.
Veröffentlicht: (2024)