ExploRLLM: Guiding Exploration in Reinforcement Learning with Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Runyu, Luijkx, Jelle, Ajanovic, Zlatan, Kober, Jens |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLM-Guided Task- and Affordance-Level Exploration in Reinforcement Learning
by: Luijkx, Jelle, et al.
Published: (2025)
by: Luijkx, Jelle, et al.
Published: (2025)
ASkDAgger: Active Skill-level Data Aggregation for Interactive Imitation Learning
by: Luijkx, Jelle, et al.
Published: (2025)
by: Luijkx, Jelle, et al.
Published: (2025)
Search Inspired Exploration in Reinforcement Learning
by: Sotirchos, Georgios, et al.
Published: (2026)
by: Sotirchos, Georgios, et al.
Published: (2026)
Mode Collapse Happens: Evaluating Critical Interactions in Joint Trajectory Prediction Models
by: Hugenholtz, Maarten, et al.
Published: (2025)
by: Hugenholtz, Maarten, et al.
Published: (2025)
Sequentially Teaching Sequential Tasks $(ST)^2$: Teaching Robots Long-horizon Manipulation Skills
by: Ajanović, Zlatan, et al.
Published: (2025)
by: Ajanović, Zlatan, et al.
Published: (2025)
Search-based versus Sampling-based Robot Motion Planning: A Comparative Study
by: Sotirchos, Georgios, et al.
Published: (2024)
by: Sotirchos, Georgios, et al.
Published: (2024)
SYMBOLIZER: Symbolic Model-free Task Planning with VLMs
by: Azirar, Sami, et al.
Published: (2026)
by: Azirar, Sami, et al.
Published: (2026)
Scenario-Based Hierarchical Reinforcement Learning for Automated Driving Decision Making
by: Abdelhamid, M. Youssef, et al.
Published: (2025)
by: Abdelhamid, M. Youssef, et al.
Published: (2025)
EAGERx: Graph-Based Framework for Sim2real Robot Learning
by: van der Heijden, Bas, et al.
Published: (2024)
by: van der Heijden, Bas, et al.
Published: (2024)
Search-Based Robot Motion Planning With Distance-Based Adaptive Motion Primitives
by: Kraljusic, Benjamin, et al.
Published: (2025)
by: Kraljusic, Benjamin, et al.
Published: (2025)
Explosive Jumping with Rigid and Articulated Soft Quadrupeds via Example Guided Reinforcement Learning
by: Apostolides, Georgios, et al.
Published: (2025)
by: Apostolides, Georgios, et al.
Published: (2025)
Impedance Primitive-augmented Hierarchical Reinforcement Learning for Sequential Tasks
by: Tahmaz, Amin Berjaoui, et al.
Published: (2025)
by: Tahmaz, Amin Berjaoui, et al.
Published: (2025)
MERLION: Marine ExploRation with Language guIded Online iNformative Visual Sampling and Enhancement
by: Thengane, Shrutika Vishal, et al.
Published: (2025)
by: Thengane, Shrutika Vishal, et al.
Published: (2025)
Novelty-based Tree-of-Thought Search for LLM Reasoning and Planning
by: Hamm, Leon, et al.
Published: (2026)
by: Hamm, Leon, et al.
Published: (2026)
Curriculum-Based Reinforcement Learning for Quadrupedal Jumping: A Reference-free Design
by: Atanassov, Vassil, et al.
Published: (2024)
by: Atanassov, Vassil, et al.
Published: (2024)
An Open-Loop Baseline for Reinforcement Learning Locomotion Tasks
by: Raffin, Antonin, et al.
Published: (2023)
by: Raffin, Antonin, et al.
Published: (2023)
Scalable Task Planning via Large Language Models and Structured World Representations
by: Pérez-Dattari, Rodrigo, et al.
Published: (2024)
by: Pérez-Dattari, Rodrigo, et al.
Published: (2024)
Scalable Task Planning via Large Language Models and Structured World Representations
by: Rodrigo Pérez‐Dattari, et al.
Published: (2025)
by: Rodrigo Pérez‐Dattari, et al.
Published: (2025)
Studying the Effect of Explicit Interaction Representations on Learning Scene-level Distributions of Human Trajectories
by: Mészáros, Anna, et al.
Published: (2025)
by: Mészáros, Anna, et al.
Published: (2025)
SLOPE: Search with Learned Optimal Pruning-based Expansion
by: Bokan, Davor, et al.
Published: (2024)
by: Bokan, Davor, et al.
Published: (2024)
PUMA: Deep Metric Imitation Learning for Stable Motion Primitives
by: Pérez-Dattari, Rodrigo, et al.
Published: (2023)
by: Pérez-Dattari, Rodrigo, et al.
Published: (2023)
Deep Reinforcement Learning-based Large-scale Robot Exploration
by: Cao, Yuhong, et al.
Published: (2024)
by: Cao, Yuhong, et al.
Published: (2024)
Do You Need a Hand? -- a Bimanual Robotic Dressing Assistance Scheme
by: Zhu, Jihong, et al.
Published: (2023)
by: Zhu, Jihong, et al.
Published: (2023)
Two-Stage Learning of Highly Dynamic Motions with Rigid and Articulated Soft Quadrupeds
by: Vezzi, Francecso, et al.
Published: (2023)
by: Vezzi, Francecso, et al.
Published: (2023)
Noise-conditioned Energy-based Annealed Rewards (NEAR): A Generative Framework for Imitation Learning from Observation
by: Diwan, Anish Abhijit, et al.
Published: (2025)
by: Diwan, Anish Abhijit, et al.
Published: (2025)
TrajFlow: Learning Distributions over Trajectories for Human Behavior Prediction
by: Mészáros, Anna, et al.
Published: (2023)
by: Mészáros, Anna, et al.
Published: (2023)
Set-Supervised Diffusion Policy: Learning Action-Chunking Diffusion through Corrections
by: Li, Zhaoting, et al.
Published: (2026)
by: Li, Zhaoting, et al.
Published: (2026)
Generalizable Motion Policies through Keypoint Parameterization and Transportation Maps
by: Franzese, Giovanni, et al.
Published: (2024)
by: Franzese, Giovanni, et al.
Published: (2024)
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback
by: Li, Zhaoting, et al.
Published: (2025)
by: Li, Zhaoting, et al.
Published: (2025)
Learning Adaptive Hydrodynamic Models Using Neural ODEs in Complex Conditions
by: Wang, Cong, et al.
Published: (2024)
by: Wang, Cong, et al.
Published: (2024)
HEADER: Hierarchical Robot Exploration via Attention-Based Deep Reinforcement Learning with Expert-Guided Reward
by: Cao, Yuhong, et al.
Published: (2025)
by: Cao, Yuhong, et al.
Published: (2025)
Privileged Reinforcement and Communication Learning for Distributed, Bandwidth-limited Multi-robot Exploration
by: Ma, Yixiao, et al.
Published: (2024)
by: Ma, Yixiao, et al.
Published: (2024)
CDE: Concept-Driven Exploration for Reinforcement Learning
by: Mao, Le, et al.
Published: (2025)
by: Mao, Le, et al.
Published: (2025)
MUKCa: Accurate and Affordable Cobot Calibration Without External Measurement Devices
by: Franzese, Giovanni, et al.
Published: (2025)
by: Franzese, Giovanni, et al.
Published: (2025)
RACP: Risk-Aware Contingency Planning with Multi-Modal Predictions
by: Mustafa, Khaled A., et al.
Published: (2024)
by: Mustafa, Khaled A., et al.
Published: (2024)
Sample-Efficient Reinforcement Learning with Temporal Logic Objectives: Leveraging the Task Specification to Guide Exploration
by: Kantaros, Yiannis, et al.
Published: (2024)
by: Kantaros, Yiannis, et al.
Published: (2024)
Task-agnostic Lifelong Robot Learning with Retrieval-based Weighted Local Adaptation
by: Yang, Pengzhi, et al.
Published: (2024)
by: Yang, Pengzhi, et al.
Published: (2024)
Large Language Model guided Deep Reinforcement Learning for Decision Making in Autonomous Driving
by: Pang, Hao, et al.
Published: (2024)
by: Pang, Hao, et al.
Published: (2024)
PrefCLM: Enhancing Preference-based Reinforcement Learning with Crowdsourced Large Language Models
by: Wang, Ruiqi, et al.
Published: (2024)
by: Wang, Ruiqi, et al.
Published: (2024)
Can Tabular Foundation Models Guide Exploration in Robot Policy Learning?
by: Ou, Buqing, et al.
Published: (2026)
by: Ou, Buqing, et al.
Published: (2026)
Similar Items
-
LLM-Guided Task- and Affordance-Level Exploration in Reinforcement Learning
by: Luijkx, Jelle, et al.
Published: (2025) -
ASkDAgger: Active Skill-level Data Aggregation for Interactive Imitation Learning
by: Luijkx, Jelle, et al.
Published: (2025) -
Search Inspired Exploration in Reinforcement Learning
by: Sotirchos, Georgios, et al.
Published: (2026) -
Mode Collapse Happens: Evaluating Critical Interactions in Joint Trajectory Prediction Models
by: Hugenholtz, Maarten, et al.
Published: (2025) -
Sequentially Teaching Sequential Tasks $(ST)^2$: Teaching Robots Long-horizon Manipulation Skills
by: Ajanović, Zlatan, et al.
Published: (2025)