A Continual Offline Reinforcement Learning Benchmark for Navigation Tasks
Fuente:
arXiv
Guardado en:
| Autores principales: | Kobanda, Anthony, Maillard, Odalric-Ambrym, Portelas, Rémy |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Hierarchical Subspaces of Policies for Continual Offline Reinforcement Learning
por: Kobanda, Anthony, et al.
Publicado: (2024)
por: Kobanda, Anthony, et al.
Publicado: (2024)
Offline Goal-Conditioned Reinforcement Learning with Projective Quasimetric Planning
por: Kobanda, Anthony, et al.
Publicado: (2025)
por: Kobanda, Anthony, et al.
Publicado: (2025)
Efficient Active Imitation Learning with Random Network Distillation
por: Biré, Emilien, et al.
Publicado: (2024)
por: Biré, Emilien, et al.
Publicado: (2024)
Leveraging priors on distribution functions for multi-arm bandits
por: Vashishtha, Sumit, et al.
Publicado: (2025)
por: Vashishtha, Sumit, et al.
Publicado: (2025)
How Hard is it to Confuse a World Model?
por: Radji, Waris, et al.
Publicado: (2025)
por: Radji, Waris, et al.
Publicado: (2025)
The regret lower bound for communicating Markov Decision Processes
por: Boone, Victor, et al.
Publicado: (2025)
por: Boone, Victor, et al.
Publicado: (2025)
The Confusing Instance Principle for Online Linear Quadratic Control
por: Radji, Waris, et al.
Publicado: (2025)
por: Radji, Waris, et al.
Publicado: (2025)
Asymptotically Optimal Problem-Dependent Bandit Policies for Transfer Learning
por: Prevost, Adrien, et al.
Publicado: (2025)
por: Prevost, Adrien, et al.
Publicado: (2025)
How to Shrink Confidence Sets for Many Equivalent Discrete Distributions?
por: Maillard, Odalric-Ambrym, et al.
Publicado: (2024)
por: Maillard, Odalric-Ambrym, et al.
Publicado: (2024)
Offline Reinforcement Learning of High-Quality Behaviors Under Robust Style Alignment
por: Petitbois, Mathieu, et al.
Publicado: (2026)
por: Petitbois, Mathieu, et al.
Publicado: (2026)
Pliable rejection sampling
por: Erraqabi, Akram, et al.
Publicado: (2026)
por: Erraqabi, Akram, et al.
Publicado: (2026)
Offline Learning of Controllable Diverse Behaviors
por: Petitbois, Mathieu, et al.
Publicado: (2025)
por: Petitbois, Mathieu, et al.
Publicado: (2025)
Provably Efficient Exploration in Reward Machines with Low Regret
por: Bourel, Hippolyte, et al.
Publicado: (2024)
por: Bourel, Hippolyte, et al.
Publicado: (2024)
Interpolation pour l'augmentation de donnees : Application à la gestion des adventices de la canne a sucre a la Reunion
por: Ferber, Frederick Fabre, et al.
Publicado: (2025)
por: Ferber, Frederick Fabre, et al.
Publicado: (2025)
Navigation with QPHIL: Quantizing Planner for Hierarchical Implicit Q-Learning
por: Canesse, Alexi, et al.
Publicado: (2024)
por: Canesse, Alexi, et al.
Publicado: (2024)
AdaStop: adaptive statistical testing for sound comparisons of Deep RL agents
por: Mathieu, Timothée, et al.
Publicado: (2023)
por: Mathieu, Timothée, et al.
Publicado: (2023)
Intrinsic-Energy Joint Embedding Predictive Architectures Induce Quasimetric Spaces
por: Kobanda, Anthony, et al.
Publicado: (2026)
por: Kobanda, Anthony, et al.
Publicado: (2026)
Power Mean Estimation in Stochastic Monte-Carlo Tree_Search
por: Dam, Tuan, et al.
Publicado: (2024)
por: Dam, Tuan, et al.
Publicado: (2024)
Using Curiosity for an Even Representation of Tasks in Continual Offline Reinforcement Learning
por: Pathmanathan, Pankayaraj, et al.
Publicado: (2023)
por: Pathmanathan, Pankayaraj, et al.
Publicado: (2023)
OER: Offline Experience Replay for Continual Offline Reinforcement Learning
por: Gai, Sibo, et al.
Publicado: (2023)
por: Gai, Sibo, et al.
Publicado: (2023)
Entropy Regularized Task Representation Learning for Offline Meta-Reinforcement Learning
por: Nakhaei, Mohammadreza, et al.
Publicado: (2024)
por: Nakhaei, Mohammadreza, et al.
Publicado: (2024)
Benchmarking Offline Multi-Objective Reinforcement Learning in Critical Care
por: Bansal, Aryaman, et al.
Publicado: (2025)
por: Bansal, Aryaman, et al.
Publicado: (2025)
Benchmarks for Reinforcement Learning with Biased Offline Data and Imperfect Simulators
por: Linial, Ori, et al.
Publicado: (2024)
por: Linial, Ori, et al.
Publicado: (2024)
Data-Incremental Continual Offline Reinforcement Learning
por: Gai, Sibo, et al.
Publicado: (2024)
por: Gai, Sibo, et al.
Publicado: (2024)
A Benchmark Environment for Offline Reinforcement Learning in Racing Games
por: Macaluso, Girolamo, et al.
Publicado: (2024)
por: Macaluso, Girolamo, et al.
Publicado: (2024)
Skills Regularized Task Decomposition for Multi-task Offline Reinforcement Learning
por: Yoo, Minjong, et al.
Publicado: (2024)
por: Yoo, Minjong, et al.
Publicado: (2024)
A Real-World Quadrupedal Locomotion Benchmark for Offline Reinforcement Learning
por: Zhang, Hongyin, et al.
Publicado: (2023)
por: Zhang, Hongyin, et al.
Publicado: (2023)
Solving Continual Offline Reinforcement Learning with Decision Transformer
por: Huang, Kaixin, et al.
Publicado: (2024)
por: Huang, Kaixin, et al.
Publicado: (2024)
Task-Aware Harmony Multi-Task Decision Transformer for Offline Reinforcement Learning
por: Fan, Ziqing, et al.
Publicado: (2024)
por: Fan, Ziqing, et al.
Publicado: (2024)
Discovering Multiple Solutions from a Single Task in Offline Reinforcement Learning
por: Osa, Takayuki, et al.
Publicado: (2024)
por: Osa, Takayuki, et al.
Publicado: (2024)
Urban-Focused Multi-Task Offline Reinforcement Learning with Contrastive Data Sharing
por: Zhao, Xinbo, et al.
Publicado: (2024)
por: Zhao, Xinbo, et al.
Publicado: (2024)
HarmoDT: Harmony Multi-Task Decision Transformer for Offline Reinforcement Learning
por: Hu, Shengchao, et al.
Publicado: (2024)
por: Hu, Shengchao, et al.
Publicado: (2024)
Model-Based Reinforcement Learning with Multi-Task Offline Pretraining
por: Pan, Minting, et al.
Publicado: (2023)
por: Pan, Minting, et al.
Publicado: (2023)
Goal-Oriented Skill Abstraction for Offline Multi-Task Reinforcement Learning
por: He, Jinmin, et al.
Publicado: (2025)
por: He, Jinmin, et al.
Publicado: (2025)
Ensemble Successor Representations for Task Generalization in Offline-to-Online Reinforcement Learning
por: Wang, Changhong, et al.
Publicado: (2024)
por: Wang, Changhong, et al.
Publicado: (2024)
Operator Models for Continuous-Time Offline Reinforcement Learning
por: Hoischen, Nicolas, et al.
Publicado: (2025)
por: Hoischen, Nicolas, et al.
Publicado: (2025)
Aquatic Navigation: A Challenging Benchmark for Deep Reinforcement Learning
por: Corsi, Davide, et al.
Publicado: (2024)
por: Corsi, Davide, et al.
Publicado: (2024)
Offline Trajectory Optimization for Offline Reinforcement Learning
por: Zhao, Ziqi, et al.
Publicado: (2024)
por: Zhao, Ziqi, et al.
Publicado: (2024)
Continual Diffuser (CoD): Mastering Continual Offline Reinforcement Learning with Experience Rehearsal
por: Hu, Jifeng, et al.
Publicado: (2024)
por: Hu, Jifeng, et al.
Publicado: (2024)
Equivariant Offline Reinforcement Learning
por: Tangri, Arsh, et al.
Publicado: (2024)
por: Tangri, Arsh, et al.
Publicado: (2024)
Ejemplares similares
-
Hierarchical Subspaces of Policies for Continual Offline Reinforcement Learning
por: Kobanda, Anthony, et al.
Publicado: (2024) -
Offline Goal-Conditioned Reinforcement Learning with Projective Quasimetric Planning
por: Kobanda, Anthony, et al.
Publicado: (2025) -
Efficient Active Imitation Learning with Random Network Distillation
por: Biré, Emilien, et al.
Publicado: (2024) -
Leveraging priors on distribution functions for multi-arm bandits
por: Vashishtha, Sumit, et al.
Publicado: (2025) -
How Hard is it to Confuse a World Model?
por: Radji, Waris, et al.
Publicado: (2025)