Explore to Generalize in Zero-Shot RL
Fuente:
arXiv
Guardado en:
| Autores principales: | Zisselman, Ev, Lavie, Itai, Soudry, Daniel, Tamar, Aviv |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Blindfolded Experts Generalize Better: Insights from Robotic Manipulation and Videogames
por: Zisselman, Ev, et al.
Publicado: (2025)
por: Zisselman, Ev, et al.
Publicado: (2025)
DDLP: Unsupervised Object-Centric Video Prediction with Deep Dynamic Latent Particles
por: Daniel, Tal, et al.
Publicado: (2023)
por: Daniel, Tal, et al.
Publicado: (2023)
Test-Time Regret Minimization in Meta Reinforcement Learning
por: Mutti, Mirco, et al.
Publicado: (2024)
por: Mutti, Mirco, et al.
Publicado: (2024)
Entity-Centric Reinforcement Learning for Object Manipulation from Pixels
por: Haramati, Dan, et al.
Publicado: (2024)
por: Haramati, Dan, et al.
Publicado: (2024)
PLUMAGE: Probabilistic Low rank Unbiased Min Variance Gradient Estimator for Efficient Large Model Training
por: Haroush, Matan, et al.
Publicado: (2025)
por: Haroush, Matan, et al.
Publicado: (2025)
Meta Reinforcement Learning with Finite Training Tasks -- a Density Estimation Approach
por: Rimon, Zohar, et al.
Publicado: (2022)
por: Rimon, Zohar, et al.
Publicado: (2022)
Grounding LTL Tasks in Sub-Symbolic RL Environments for Zero-Shot Generalization
por: Pannacci, Matteo, et al.
Publicado: (2026)
por: Pannacci, Matteo, et al.
Publicado: (2026)
A Classification View on Meta Learning Bandits
por: Mutti, Mirco, et al.
Publicado: (2025)
por: Mutti, Mirco, et al.
Publicado: (2025)
RoboArm-NMP: a Learning Environment for Neural Motion Planning
por: Jurgenson, Tom, et al.
Publicado: (2024)
por: Jurgenson, Tom, et al.
Publicado: (2024)
Assessing the Zero-Shot Capabilities of LLMs for Action Evaluation in RL
por: Pignatelli, Eduardo, et al.
Publicado: (2024)
por: Pignatelli, Eduardo, et al.
Publicado: (2024)
TGRL: An Algorithm for Teacher Guided Reinforcement Learning
por: Shenfeld, Idan, et al.
Publicado: (2023)
por: Shenfeld, Idan, et al.
Publicado: (2023)
Zero-Shot Generalization of Vision-Based RL Without Data Augmentation
por: Batra, Sumeet, et al.
Publicado: (2024)
por: Batra, Sumeet, et al.
Publicado: (2024)
Zero-Shot Instruction Following in RL via Structured LTL Representations
por: Giuri, Mattia, et al.
Publicado: (2025)
por: Giuri, Mattia, et al.
Publicado: (2025)
Zero-Shot Instruction Following in RL via Structured LTL Representations
por: Jackermeier, Mathias, et al.
Publicado: (2026)
por: Jackermeier, Mathias, et al.
Publicado: (2026)
Temporal Difference Calibration in Sequential Tasks: Application to Vision-Language-Action Models
por: Francis-Meretzki, Shelly, et al.
Publicado: (2026)
por: Francis-Meretzki, Shelly, et al.
Publicado: (2026)
Temperature is All You Need for Generalization in Langevin Dynamics and other Markov Processes
por: Harel, Itamar, et al.
Publicado: (2025)
por: Harel, Itamar, et al.
Publicado: (2025)
Toward Artificial Palpation: Representation Learning of Touch on Soft Bodies
por: Rimon, Zohar, et al.
Publicado: (2025)
por: Rimon, Zohar, et al.
Publicado: (2025)
MAMBA: an Effective World Model Approach for Meta-Reinforcement Learning
por: Rimon, Zohar, et al.
Publicado: (2024)
por: Rimon, Zohar, et al.
Publicado: (2024)
Few-Shot Inspired Generative Zero-Shot Learning
por: Shohag, Md Shakil Ahamed, et al.
Publicado: (2025)
por: Shohag, Md Shakil Ahamed, et al.
Publicado: (2025)
The Implicit Bias of Gradient Descent on Separable Multiclass Data
por: Ravi, Hrithik, et al.
Publicado: (2024)
por: Ravi, Hrithik, et al.
Publicado: (2024)
Foldable SuperNets: Scalable Merging of Transformers with Different Initializations and Tasks
por: Kinderman, Edan, et al.
Publicado: (2024)
por: Kinderman, Edan, et al.
Publicado: (2024)
Zero-Shot LLMs in Human-in-the-Loop RL: Replacing Human Feedback for Reward Shaping
por: Nazir, Mohammad Saif, et al.
Publicado: (2025)
por: Nazir, Mohammad Saif, et al.
Publicado: (2025)
Evolution Strategies for Deep RL pretraining
por: Martínez, Adrian, et al.
Publicado: (2026)
por: Martínez, Adrian, et al.
Publicado: (2026)
A Generalization Theory for Zero-Shot Prediction
por: Mehta, Ronak, et al.
Publicado: (2025)
por: Mehta, Ronak, et al.
Publicado: (2025)
Towards Cheaper Inference in Deep Networks with Lower Bit-Width Accumulators
por: Blumenfeld, Yaniv, et al.
Publicado: (2024)
por: Blumenfeld, Yaniv, et al.
Publicado: (2024)
The Joint Effect of Task Similarity and Overparameterization on Catastrophic Forgetting -- An Analytical Model
por: Goldfarb, Daniel, et al.
Publicado: (2024)
por: Goldfarb, Daniel, et al.
Publicado: (2024)
Zero-Shot Robustification of Zero-Shot Models
por: Adila, Dyah, et al.
Publicado: (2023)
por: Adila, Dyah, et al.
Publicado: (2023)
Stable Minima Cannot Overfit in Univariate ReLU Networks: Generalization by Large Step Sizes
por: Qiao, Dan, et al.
Publicado: (2024)
por: Qiao, Dan, et al.
Publicado: (2024)
Minimum Variance Unbiased N:M Sparsity for the Neural Gradients
por: Chmiel, Brian, et al.
Publicado: (2022)
por: Chmiel, Brian, et al.
Publicado: (2022)
FP4 All the Way: Fully Quantized Training of LLMs
por: Chmiel, Brian, et al.
Publicado: (2025)
por: Chmiel, Brian, et al.
Publicado: (2025)
Scaling FP8 training to trillion-token LLMs
por: Fishman, Maxim, et al.
Publicado: (2024)
por: Fishman, Maxim, et al.
Publicado: (2024)
How Uniform Random Weights Induce Non-uniform Bias: Typical Interpolating Neural Networks Generalize with Narrow Teachers
por: Buzaglo, Gon, et al.
Publicado: (2024)
por: Buzaglo, Gon, et al.
Publicado: (2024)
Latent Particle World Models: Self-supervised Object-centric Stochastic Dynamics Modeling
por: Daniel, Tal, et al.
Publicado: (2026)
por: Daniel, Tal, et al.
Publicado: (2026)
How do Minimum-Norm Shallow Denoisers Look in Function Space?
por: Zeno, Chen, et al.
Publicado: (2023)
por: Zeno, Chen, et al.
Publicado: (2023)
Tensor-Parallelism with Partially Synchronized Activations
por: Lamprecht, Itay, et al.
Publicado: (2025)
por: Lamprecht, Itay, et al.
Publicado: (2025)
Few-Shot Task Learning through Inverse Generative Modeling
por: Netanyahu, Aviv, et al.
Publicado: (2024)
por: Netanyahu, Aviv, et al.
Publicado: (2024)
Learning to Route Among Specialized Experts for Zero-Shot Generalization
por: Muqeeth, Mohammed, et al.
Publicado: (2024)
por: Muqeeth, Mohammed, et al.
Publicado: (2024)
Hierarchical Entity-centric Reinforcement Learning with Factored Subgoal Diffusion
por: Haramati, Dan, et al.
Publicado: (2026)
por: Haramati, Dan, et al.
Publicado: (2026)
What Matters in Learning A Zero-Shot Sim-to-Real RL Policy for Quadrotor Control? A Comprehensive Study
por: Chen, Jiayu, et al.
Publicado: (2024)
por: Chen, Jiayu, et al.
Publicado: (2024)
Prompt-Tuning Bandits: Enabling Few-Shot Generalization for Efficient Multi-Task Offline RL
por: Rietz, Finn, et al.
Publicado: (2025)
por: Rietz, Finn, et al.
Publicado: (2025)
Ejemplares similares
-
Blindfolded Experts Generalize Better: Insights from Robotic Manipulation and Videogames
por: Zisselman, Ev, et al.
Publicado: (2025) -
DDLP: Unsupervised Object-Centric Video Prediction with Deep Dynamic Latent Particles
por: Daniel, Tal, et al.
Publicado: (2023) -
Test-Time Regret Minimization in Meta Reinforcement Learning
por: Mutti, Mirco, et al.
Publicado: (2024) -
Entity-Centric Reinforcement Learning for Object Manipulation from Pixels
por: Haramati, Dan, et al.
Publicado: (2024) -
PLUMAGE: Probabilistic Low rank Unbiased Min Variance Gradient Estimator for Efficient Large Model Training
por: Haroush, Matan, et al.
Publicado: (2025)