A Tutorial on Meta-Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Beck, Jacob, Vuorio, Risto, Liu, Evan Zheran, Xiong, Zheng, Zintgraf, Luisa, Finn, Chelsea, Whiteson, Shimon |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SplAgger: Split Aggregation for Meta-Reinforcement Learning
by: Beck, Jacob, et al.
Published: (2024)
by: Beck, Jacob, et al.
Published: (2024)
Distilling Morphology-Conditioned Hypernetworks for Efficient Universal Morphology Control
by: Xiong, Zheng, et al.
Published: (2024)
by: Xiong, Zheng, et al.
Published: (2024)
IGDrivSim: A Benchmark for the Imitation Gap in Autonomous Driving
by: Grislain, Clémence, et al.
Published: (2024)
by: Grislain, Clémence, et al.
Published: (2024)
A Bayesian Solution To The Imitation Gap
by: Vuorio, Risto, et al.
Published: (2024)
by: Vuorio, Risto, et al.
Published: (2024)
A Survey of In-Context Reinforcement Learning
by: Moeini, Amir, et al.
Published: (2025)
by: Moeini, Amir, et al.
Published: (2025)
How Should We Meta-Learn Reinforcement Learning Algorithms?
by: Goldie, Alexander David, et al.
Published: (2025)
by: Goldie, Alexander David, et al.
Published: (2025)
GoalLadder: Incremental Goal Discovery with Vision-Language Models
by: Zakharov, Alexey, et al.
Published: (2025)
by: Zakharov, Alexey, et al.
Published: (2025)
A Clean Slate for Offline Reinforcement Learning
by: Jackson, Matthew Thomas, et al.
Published: (2025)
by: Jackson, Matthew Thomas, et al.
Published: (2025)
Scalable Meta-Learning via Mixed-Mode Differentiation
by: Kemaev, Iurii, et al.
Published: (2025)
by: Kemaev, Iurii, et al.
Published: (2025)
Can Learned Optimization Make Reinforcement Learning Less Difficult?
by: Goldie, Alexander David, et al.
Published: (2024)
by: Goldie, Alexander David, et al.
Published: (2024)
HyperVLA: Efficient Inference in Vision-Language-Action Models via Hypernetworks
by: Xiong, Zheng, et al.
Published: (2025)
by: Xiong, Zheng, et al.
Published: (2025)
Discovering Temporally-Aware Reinforcement Learning Algorithms
by: Jackson, Matthew Thomas, et al.
Published: (2024)
by: Jackson, Matthew Thomas, et al.
Published: (2024)
Deconfounding Imitation Learning with Variational Inference
by: Vuorio, Risto, et al.
Published: (2022)
by: Vuorio, Risto, et al.
Published: (2022)
Rate-Informed Discovery via Bayesian Adaptive Multifidelity Sampling
by: Sinha, Aman, et al.
Published: (2024)
by: Sinha, Aman, et al.
Published: (2024)
EXPO: Stable Reinforcement Learning with Expressive Policies
by: Dong, Perry, et al.
Published: (2025)
by: Dong, Perry, et al.
Published: (2025)
Bayesian Exploration Networks
by: Fellows, Mattie, et al.
Published: (2023)
by: Fellows, Mattie, et al.
Published: (2023)
Grounding by Trying: LLMs with Reinforcement Learning-Enhanced Retrieval
by: Hsu, Sheryl, et al.
Published: (2024)
by: Hsu, Sheryl, et al.
Published: (2024)
Reinforcement Learning via Implicit Imitation Guidance
by: Dong, Perry, et al.
Published: (2025)
by: Dong, Perry, et al.
Published: (2025)
Polychromic Objectives for Reinforcement Learning
by: Hamid, Jubayer Ibn, et al.
Published: (2025)
by: Hamid, Jubayer Ibn, et al.
Published: (2025)
Action-Constrained Imitation Learning
by: Yeh, Chia-Han, et al.
Published: (2025)
by: Yeh, Chia-Han, et al.
Published: (2025)
Affordance-Guided Reinforcement Learning via Visual Prompting
by: Lee, Olivia Y., et al.
Published: (2024)
by: Lee, Olivia Y., et al.
Published: (2024)
Equivariant Networks for Zero-Shot Coordination
by: Muglich, Darius, et al.
Published: (2022)
by: Muglich, Darius, et al.
Published: (2022)
A Tutorial: An Intuitive Explanation of Offline Reinforcement Learning Theory
by: Che, Fengdi
Published: (2025)
by: Che, Fengdi
Published: (2025)
Feedback Descent: Open-Ended Text Optimization via Pairwise Comparison
by: Lee, Yoonho, et al.
Published: (2025)
by: Lee, Yoonho, et al.
Published: (2025)
Learning Long-Context Diffusion Policies via Past-Token Prediction
by: Torne, Marcel, et al.
Published: (2025)
by: Torne, Marcel, et al.
Published: (2025)
UniGen: Unified Modeling of Initial Agent States and Trajectories for Generating Autonomous Driving Scenarios
by: Mahjourian, Reza, et al.
Published: (2024)
by: Mahjourian, Reza, et al.
Published: (2024)
Learning to Drive in New Cities Without Human Demonstrations
by: Wang, Zilin, et al.
Published: (2026)
by: Wang, Zilin, et al.
Published: (2026)
AutoBencher: Towards Declarative Benchmark Construction
by: Li, Xiang Lisa, et al.
Published: (2024)
by: Li, Xiang Lisa, et al.
Published: (2024)
DataRater: Meta-Learned Dataset Curation
by: Calian, Dan A., et al.
Published: (2025)
by: Calian, Dan A., et al.
Published: (2025)
Just Enough Thinking: Efficient Reasoning with Adaptive Length Penalties Reinforcement Learning
by: Xiang, Violet, et al.
Published: (2025)
by: Xiang, Violet, et al.
Published: (2025)
Efficient Imitation Learning with Conservative World Models
by: Kolev, Victor, et al.
Published: (2024)
by: Kolev, Victor, et al.
Published: (2024)
Self-Guided Masked Autoencoders for Domain-Agnostic Self-Supervised Learning
by: Xie, Johnathan, et al.
Published: (2024)
by: Xie, Johnathan, et al.
Published: (2024)
Policy-Guided Diffusion
by: Jackson, Matthew Thomas, et al.
Published: (2024)
by: Jackson, Matthew Thomas, et al.
Published: (2024)
Universal Neural Functionals
by: Zhou, Allan, et al.
Published: (2024)
by: Zhou, Allan, et al.
Published: (2024)
Adam on Local Time: Addressing Nonstationarity in RL with Relative Adam Timesteps
by: Ellis, Benjamin, et al.
Published: (2024)
by: Ellis, Benjamin, et al.
Published: (2024)
From $r$ to $Q^*$: Your Language Model is Secretly a Q-Function
by: Rafailov, Rafael, et al.
Published: (2024)
by: Rafailov, Rafael, et al.
Published: (2024)
Value Flows
by: Dong, Perry, et al.
Published: (2025)
by: Dong, Perry, et al.
Published: (2025)
Metalic: Meta-Learning In-Context with Protein Language Models
by: Beck, Jacob, et al.
Published: (2024)
by: Beck, Jacob, et al.
Published: (2024)
Offline RLAIF: Piloting VLM Feedback for RL via SFO
by: Beck, Jacob
Published: (2025)
by: Beck, Jacob
Published: (2025)
Model-Based Reinforcement Learning for Atari
by: Kaiser, Lukasz, et al.
Published: (2019)
by: Kaiser, Lukasz, et al.
Published: (2019)
Similar Items
-
SplAgger: Split Aggregation for Meta-Reinforcement Learning
by: Beck, Jacob, et al.
Published: (2024) -
Distilling Morphology-Conditioned Hypernetworks for Efficient Universal Morphology Control
by: Xiong, Zheng, et al.
Published: (2024) -
IGDrivSim: A Benchmark for the Imitation Gap in Autonomous Driving
by: Grislain, Clémence, et al.
Published: (2024) -
A Bayesian Solution To The Imitation Gap
by: Vuorio, Risto, et al.
Published: (2024) -
A Survey of In-Context Reinforcement Learning
by: Moeini, Amir, et al.
Published: (2025)