Improving Zero-Shot Offline RL via Behavioral Task Sampling
Fuente:
arXiv
Saved in:
| Main Authors: | Bendib, Nazim, Perrin-Gilbert, Nicolas, Sigaud, Olivier |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Single-Reset Divide & Conquer Imitation Learning
by: Chenu, Alexandre, et al.
Published: (2024)
by: Chenu, Alexandre, et al.
Published: (2024)
AFU: Actor-Free critic Updates in off-policy RL for continuous control
by: Perrin-Gilbert, Nicolas
Published: (2024)
by: Perrin-Gilbert, Nicolas
Published: (2024)
Imagine Beyond! Distributionally Robust Auto-Encoding for State Space Coverage in Online Reinforcement Learning
by: Castanet, Nicolas, et al.
Published: (2025)
by: Castanet, Nicolas, et al.
Published: (2025)
A Definition of Open-Ended Learning Problems for Goal-Conditioned Agents
by: Sigaud, Olivier, et al.
Published: (2023)
by: Sigaud, Olivier, et al.
Published: (2023)
CoViews: Adaptive Augmentation Using Cooperative Views for Enhanced Contrastive Learning
by: Bendib, Nazim
Published: (2024)
by: Bendib, Nazim
Published: (2024)
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
by: Gaven, Loris, et al.
Published: (2024)
by: Gaven, Loris, et al.
Published: (2024)
Physics-Informed Model and Hybrid Planning for Efficient Dyna-Style Reinforcement Learning
by: Asri, Zakariae El, et al.
Published: (2024)
by: Asri, Zakariae El, et al.
Published: (2024)
From Goal-Conditioned to Language-Conditioned Agents via Vision-Language Models
by: Cachet, Theo, et al.
Published: (2024)
by: Cachet, Theo, et al.
Published: (2024)
Grounding LTL Tasks in Sub-Symbolic RL Environments for Zero-Shot Generalization
by: Pannacci, Matteo, et al.
Published: (2026)
by: Pannacci, Matteo, et al.
Published: (2026)
Guided Flow Policy: Learning from High-Value Actions in Offline Reinforcement Learning
by: Tiofack, Franki Nguimatsia, et al.
Published: (2025)
by: Tiofack, Franki Nguimatsia, et al.
Published: (2025)
A tale of two goals: leveraging sequentiality in multi-goal scenarios
by: Serris, Olivier, et al.
Published: (2025)
by: Serris, Olivier, et al.
Published: (2025)
Zero-Shot Instruction Following in RL via Structured LTL Representations
by: Jackermeier, Mathias, et al.
Published: (2026)
by: Jackermeier, Mathias, et al.
Published: (2026)
Zero-Shot Instruction Following in RL via Structured LTL Representations
by: Giuri, Mattia, et al.
Published: (2025)
by: Giuri, Mattia, et al.
Published: (2025)
Provable Zero-Shot Generalization in Offline Reinforcement Learning
by: Wang, Zhiyong, et al.
Published: (2025)
by: Wang, Zhiyong, et al.
Published: (2025)
Diffusion-Based Offline RL for Improved Decision-Making in Augmented ARC Task
by: Kim, Yunho, et al.
Published: (2024)
by: Kim, Yunho, et al.
Published: (2024)
Zero-Shot Statistical Downscaling via Diffusion Posterior Sampling
by: Tie, Ruian, et al.
Published: (2026)
by: Tie, Ruian, et al.
Published: (2026)
Assessing the Zero-Shot Capabilities of LLMs for Action Evaluation in RL
by: Pignatelli, Eduardo, et al.
Published: (2024)
by: Pignatelli, Eduardo, et al.
Published: (2024)
Inferring Behavior-Specific Context Improves Zero-Shot Generalization in Reinforcement Learning
by: Ndir, Tidiane Camaret, et al.
Published: (2024)
by: Ndir, Tidiane Camaret, et al.
Published: (2024)
Scaling Offline RL via Efficient and Expressive Shortcut Models
by: Espinosa-Dice, Nicolas, et al.
Published: (2025)
by: Espinosa-Dice, Nicolas, et al.
Published: (2025)
Budgeting Counterfactual for Offline RL
by: Liu, Yao, et al.
Published: (2023)
by: Liu, Yao, et al.
Published: (2023)
Bus-Conditioned Zero-Shot Trajectory Generation via Task Arithmetic
by: Liu, Shuai, et al.
Published: (2026)
by: Liu, Shuai, et al.
Published: (2026)
Offline RLAIF: Piloting VLM Feedback for RL via SFO
by: Beck, Jacob
Published: (2025)
by: Beck, Jacob
Published: (2025)
Selective Uncertainty Propagation in Offline RL
by: Krishnamurthy, Sanath Kumar, et al.
Published: (2023)
by: Krishnamurthy, Sanath Kumar, et al.
Published: (2023)
Decoupled Prioritized Resampling for Offline RL
by: Yue, Yang, et al.
Published: (2023)
by: Yue, Yang, et al.
Published: (2023)
Augmenting Offline RL with Unlabeled Data
by: Wang, Zhao, et al.
Published: (2024)
by: Wang, Zhao, et al.
Published: (2024)
Autotelic Agents with Intrinsically Motivated Goal-Conditioned Reinforcement Learning: a Short Survey
by: Colas, Cédric, et al.
Published: (2020)
by: Colas, Cédric, et al.
Published: (2020)
PRISM: Perception Reasoning Interleaved for Sequential Decision Making
by: Aissi, Mohamed Salim, et al.
Published: (2026)
by: Aissi, Mohamed Salim, et al.
Published: (2026)
An Empirical Study on the Effectiveness of Incorporating Offline RL As Online RL Subroutines
by: Su, Jianhai, et al.
Published: (2025)
by: Su, Jianhai, et al.
Published: (2025)
Action-Free Offline-to-Online RL via Discretised State Policies
by: Neggatu, Natinael Solomon, et al.
Published: (2026)
by: Neggatu, Natinael Solomon, et al.
Published: (2026)
CURIOUS: Intrinsically Motivated Modular Multi-Goal Reinforcement Learning
by: Colas, Cédric, et al.
Published: (2018)
by: Colas, Cédric, et al.
Published: (2018)
Explicit Logic Channel for Validation and Enhancement of MLLMs on Zero-Shot Tasks
by: Leong, Mei Chee, et al.
Published: (2026)
by: Leong, Mei Chee, et al.
Published: (2026)
Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone
by: Mark, Max Sobol, et al.
Published: (2024)
by: Mark, Max Sobol, et al.
Published: (2024)
Frog Soup: Zero-Shot, In-Context, and Sample-Efficient Frogger Agents
by: Li, Xiang, et al.
Published: (2025)
by: Li, Xiang, et al.
Published: (2025)
Find the Fruit: Zero-Shot Sim2Real RL for Occlusion-Aware Plant Manipulation
by: Subedi, Nitesh, et al.
Published: (2025)
by: Subedi, Nitesh, et al.
Published: (2025)
General Flexible $f$-divergence for Challenging Offline RL Datasets with Low Stochasticity and Diverse Behavior Policies
by: Wang, Jianxun, et al.
Published: (2026)
by: Wang, Jianxun, et al.
Published: (2026)
A Tractable Inference Perspective of Offline RL
by: Liu, Xuejie, et al.
Published: (2023)
by: Liu, Xuejie, et al.
Published: (2023)
Design Considerations in Offline Preference-based RL
by: Agarwal, Alekh, et al.
Published: (2025)
by: Agarwal, Alekh, et al.
Published: (2025)
Are Expressive Models Truly Necessary for Offline RL?
by: Wang, Guan, et al.
Published: (2024)
by: Wang, Guan, et al.
Published: (2024)
OGBench: Benchmarking Offline Goal-Conditioned RL
by: Park, Seohong, et al.
Published: (2024)
by: Park, Seohong, et al.
Published: (2024)
Offline-Boosted Actor-Critic: Adaptively Blending Optimal Historical Behaviors in Deep Off-Policy RL
by: Luo, Yu, et al.
Published: (2024)
by: Luo, Yu, et al.
Published: (2024)
Similar Items
-
Single-Reset Divide & Conquer Imitation Learning
by: Chenu, Alexandre, et al.
Published: (2024) -
AFU: Actor-Free critic Updates in off-policy RL for continuous control
by: Perrin-Gilbert, Nicolas
Published: (2024) -
Imagine Beyond! Distributionally Robust Auto-Encoding for State Space Coverage in Online Reinforcement Learning
by: Castanet, Nicolas, et al.
Published: (2025) -
A Definition of Open-Ended Learning Problems for Goal-Conditioned Agents
by: Sigaud, Olivier, et al.
Published: (2023) -
CoViews: Adaptive Augmentation Using Cooperative Views for Enhanced Contrastive Learning
by: Bendib, Nazim
Published: (2024)