Explore to Generalize in Zero-Shot RL
Fuente:
arXiv
Saved in:
| Main Authors: | Zisselman, Ev, Lavie, Itai, Soudry, Daniel, Tamar, Aviv |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Blindfolded Experts Generalize Better: Insights from Robotic Manipulation and Videogames
by: Zisselman, Ev, et al.
Published: (2025)
by: Zisselman, Ev, et al.
Published: (2025)
DDLP: Unsupervised Object-Centric Video Prediction with Deep Dynamic Latent Particles
by: Daniel, Tal, et al.
Published: (2023)
by: Daniel, Tal, et al.
Published: (2023)
Test-Time Regret Minimization in Meta Reinforcement Learning
by: Mutti, Mirco, et al.
Published: (2024)
by: Mutti, Mirco, et al.
Published: (2024)
Entity-Centric Reinforcement Learning for Object Manipulation from Pixels
by: Haramati, Dan, et al.
Published: (2024)
by: Haramati, Dan, et al.
Published: (2024)
PLUMAGE: Probabilistic Low rank Unbiased Min Variance Gradient Estimator for Efficient Large Model Training
by: Haroush, Matan, et al.
Published: (2025)
by: Haroush, Matan, et al.
Published: (2025)
Meta Reinforcement Learning with Finite Training Tasks -- a Density Estimation Approach
by: Rimon, Zohar, et al.
Published: (2022)
by: Rimon, Zohar, et al.
Published: (2022)
Grounding LTL Tasks in Sub-Symbolic RL Environments for Zero-Shot Generalization
by: Pannacci, Matteo, et al.
Published: (2026)
by: Pannacci, Matteo, et al.
Published: (2026)
A Classification View on Meta Learning Bandits
by: Mutti, Mirco, et al.
Published: (2025)
by: Mutti, Mirco, et al.
Published: (2025)
RoboArm-NMP: a Learning Environment for Neural Motion Planning
by: Jurgenson, Tom, et al.
Published: (2024)
by: Jurgenson, Tom, et al.
Published: (2024)
Assessing the Zero-Shot Capabilities of LLMs for Action Evaluation in RL
by: Pignatelli, Eduardo, et al.
Published: (2024)
by: Pignatelli, Eduardo, et al.
Published: (2024)
TGRL: An Algorithm for Teacher Guided Reinforcement Learning
by: Shenfeld, Idan, et al.
Published: (2023)
by: Shenfeld, Idan, et al.
Published: (2023)
Zero-Shot Generalization of Vision-Based RL Without Data Augmentation
by: Batra, Sumeet, et al.
Published: (2024)
by: Batra, Sumeet, et al.
Published: (2024)
Zero-Shot Instruction Following in RL via Structured LTL Representations
by: Giuri, Mattia, et al.
Published: (2025)
by: Giuri, Mattia, et al.
Published: (2025)
Zero-Shot Instruction Following in RL via Structured LTL Representations
by: Jackermeier, Mathias, et al.
Published: (2026)
by: Jackermeier, Mathias, et al.
Published: (2026)
Temporal Difference Calibration in Sequential Tasks: Application to Vision-Language-Action Models
by: Francis-Meretzki, Shelly, et al.
Published: (2026)
by: Francis-Meretzki, Shelly, et al.
Published: (2026)
Temperature is All You Need for Generalization in Langevin Dynamics and other Markov Processes
by: Harel, Itamar, et al.
Published: (2025)
by: Harel, Itamar, et al.
Published: (2025)
Toward Artificial Palpation: Representation Learning of Touch on Soft Bodies
by: Rimon, Zohar, et al.
Published: (2025)
by: Rimon, Zohar, et al.
Published: (2025)
MAMBA: an Effective World Model Approach for Meta-Reinforcement Learning
by: Rimon, Zohar, et al.
Published: (2024)
by: Rimon, Zohar, et al.
Published: (2024)
Few-Shot Inspired Generative Zero-Shot Learning
by: Shohag, Md Shakil Ahamed, et al.
Published: (2025)
by: Shohag, Md Shakil Ahamed, et al.
Published: (2025)
The Implicit Bias of Gradient Descent on Separable Multiclass Data
by: Ravi, Hrithik, et al.
Published: (2024)
by: Ravi, Hrithik, et al.
Published: (2024)
Foldable SuperNets: Scalable Merging of Transformers with Different Initializations and Tasks
by: Kinderman, Edan, et al.
Published: (2024)
by: Kinderman, Edan, et al.
Published: (2024)
Zero-Shot LLMs in Human-in-the-Loop RL: Replacing Human Feedback for Reward Shaping
by: Nazir, Mohammad Saif, et al.
Published: (2025)
by: Nazir, Mohammad Saif, et al.
Published: (2025)
Evolution Strategies for Deep RL pretraining
by: Martínez, Adrian, et al.
Published: (2026)
by: Martínez, Adrian, et al.
Published: (2026)
A Generalization Theory for Zero-Shot Prediction
by: Mehta, Ronak, et al.
Published: (2025)
by: Mehta, Ronak, et al.
Published: (2025)
Towards Cheaper Inference in Deep Networks with Lower Bit-Width Accumulators
by: Blumenfeld, Yaniv, et al.
Published: (2024)
by: Blumenfeld, Yaniv, et al.
Published: (2024)
The Joint Effect of Task Similarity and Overparameterization on Catastrophic Forgetting -- An Analytical Model
by: Goldfarb, Daniel, et al.
Published: (2024)
by: Goldfarb, Daniel, et al.
Published: (2024)
Zero-Shot Robustification of Zero-Shot Models
by: Adila, Dyah, et al.
Published: (2023)
by: Adila, Dyah, et al.
Published: (2023)
Stable Minima Cannot Overfit in Univariate ReLU Networks: Generalization by Large Step Sizes
by: Qiao, Dan, et al.
Published: (2024)
by: Qiao, Dan, et al.
Published: (2024)
Minimum Variance Unbiased N:M Sparsity for the Neural Gradients
by: Chmiel, Brian, et al.
Published: (2022)
by: Chmiel, Brian, et al.
Published: (2022)
FP4 All the Way: Fully Quantized Training of LLMs
by: Chmiel, Brian, et al.
Published: (2025)
by: Chmiel, Brian, et al.
Published: (2025)
Scaling FP8 training to trillion-token LLMs
by: Fishman, Maxim, et al.
Published: (2024)
by: Fishman, Maxim, et al.
Published: (2024)
How Uniform Random Weights Induce Non-uniform Bias: Typical Interpolating Neural Networks Generalize with Narrow Teachers
by: Buzaglo, Gon, et al.
Published: (2024)
by: Buzaglo, Gon, et al.
Published: (2024)
Latent Particle World Models: Self-supervised Object-centric Stochastic Dynamics Modeling
by: Daniel, Tal, et al.
Published: (2026)
by: Daniel, Tal, et al.
Published: (2026)
How do Minimum-Norm Shallow Denoisers Look in Function Space?
by: Zeno, Chen, et al.
Published: (2023)
by: Zeno, Chen, et al.
Published: (2023)
Tensor-Parallelism with Partially Synchronized Activations
by: Lamprecht, Itay, et al.
Published: (2025)
by: Lamprecht, Itay, et al.
Published: (2025)
Few-Shot Task Learning through Inverse Generative Modeling
by: Netanyahu, Aviv, et al.
Published: (2024)
by: Netanyahu, Aviv, et al.
Published: (2024)
Learning to Route Among Specialized Experts for Zero-Shot Generalization
by: Muqeeth, Mohammed, et al.
Published: (2024)
by: Muqeeth, Mohammed, et al.
Published: (2024)
Hierarchical Entity-centric Reinforcement Learning with Factored Subgoal Diffusion
by: Haramati, Dan, et al.
Published: (2026)
by: Haramati, Dan, et al.
Published: (2026)
What Matters in Learning A Zero-Shot Sim-to-Real RL Policy for Quadrotor Control? A Comprehensive Study
by: Chen, Jiayu, et al.
Published: (2024)
by: Chen, Jiayu, et al.
Published: (2024)
Prompt-Tuning Bandits: Enabling Few-Shot Generalization for Efficient Multi-Task Offline RL
by: Rietz, Finn, et al.
Published: (2025)
by: Rietz, Finn, et al.
Published: (2025)
Similar Items
-
Blindfolded Experts Generalize Better: Insights from Robotic Manipulation and Videogames
by: Zisselman, Ev, et al.
Published: (2025) -
DDLP: Unsupervised Object-Centric Video Prediction with Deep Dynamic Latent Particles
by: Daniel, Tal, et al.
Published: (2023) -
Test-Time Regret Minimization in Meta Reinforcement Learning
by: Mutti, Mirco, et al.
Published: (2024) -
Entity-Centric Reinforcement Learning for Object Manipulation from Pixels
by: Haramati, Dan, et al.
Published: (2024) -
PLUMAGE: Probabilistic Low rank Unbiased Min Variance Gradient Estimator for Efficient Large Model Training
by: Haroush, Matan, et al.
Published: (2025)