Offline Goal-Conditioned Reinforcement Learning with Projective Quasimetric Planning
Fuente:
arXiv
Saved in:
| Main Authors: | Kobanda, Anthony, Radji, Waris, Petitbois, Mathieu, Maillard, Odalric-Ambrym, Portelas, Rémy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Continual Offline Reinforcement Learning Benchmark for Navigation Tasks
by: Kobanda, Anthony, et al.
Published: (2025)
by: Kobanda, Anthony, et al.
Published: (2025)
Hierarchical Subspaces of Policies for Continual Offline Reinforcement Learning
by: Kobanda, Anthony, et al.
Published: (2024)
by: Kobanda, Anthony, et al.
Published: (2024)
How Hard is it to Confuse a World Model?
by: Radji, Waris, et al.
Published: (2025)
by: Radji, Waris, et al.
Published: (2025)
The Confusing Instance Principle for Online Linear Quadratic Control
by: Radji, Waris, et al.
Published: (2025)
by: Radji, Waris, et al.
Published: (2025)
Intrinsic-Energy Joint Embedding Predictive Architectures Induce Quasimetric Spaces
by: Kobanda, Anthony, et al.
Published: (2026)
by: Kobanda, Anthony, et al.
Published: (2026)
Offline Reinforcement Learning of High-Quality Behaviors Under Robust Style Alignment
by: Petitbois, Mathieu, et al.
Published: (2026)
by: Petitbois, Mathieu, et al.
Published: (2026)
Offline Learning of Controllable Diverse Behaviors
by: Petitbois, Mathieu, et al.
Published: (2025)
by: Petitbois, Mathieu, et al.
Published: (2025)
Asymptotically Optimal Problem-Dependent Bandit Policies for Transfer Learning
by: Prevost, Adrien, et al.
Published: (2025)
by: Prevost, Adrien, et al.
Published: (2025)
Efficient Active Imitation Learning with Random Network Distillation
by: Biré, Emilien, et al.
Published: (2024)
by: Biré, Emilien, et al.
Published: (2024)
Leveraging priors on distribution functions for multi-arm bandits
by: Vashishtha, Sumit, et al.
Published: (2025)
by: Vashishtha, Sumit, et al.
Published: (2025)
The regret lower bound for communicating Markov Decision Processes
by: Boone, Victor, et al.
Published: (2025)
by: Boone, Victor, et al.
Published: (2025)
Navigation with QPHIL: Quantizing Planner for Hierarchical Implicit Q-Learning
by: Canesse, Alexi, et al.
Published: (2024)
by: Canesse, Alexi, et al.
Published: (2024)
How to Shrink Confidence Sets for Many Equivalent Discrete Distributions?
by: Maillard, Odalric-Ambrym, et al.
Published: (2024)
by: Maillard, Odalric-Ambrym, et al.
Published: (2024)
Offline Goal-conditioned Reinforcement Learning with Quasimetric Representations
by: Myers, Vivek, et al.
Published: (2025)
by: Myers, Vivek, et al.
Published: (2025)
Pliable rejection sampling
by: Erraqabi, Akram, et al.
Published: (2026)
by: Erraqabi, Akram, et al.
Published: (2026)
Octax: Accelerated CHIP-8 Arcade Environments for Reinforcement Learning in JAX
by: Radji, Waris, et al.
Published: (2025)
by: Radji, Waris, et al.
Published: (2025)
Multistep Quasimetric Learning for Scalable Goal-conditioned Reinforcement Learning
by: Zheng, Bill Chunyuan, et al.
Published: (2025)
by: Zheng, Bill Chunyuan, et al.
Published: (2025)
Provably Efficient Exploration in Reward Machines with Low Regret
by: Bourel, Hippolyte, et al.
Published: (2024)
by: Bourel, Hippolyte, et al.
Published: (2024)
Goal Reaching with Eikonal-Constrained Hierarchical Quasimetric Reinforcement Learning
by: Giammarino, Vittorio, et al.
Published: (2025)
by: Giammarino, Vittorio, et al.
Published: (2025)
AdaStop: adaptive statistical testing for sound comparisons of Deep RL agents
by: Mathieu, Timothée, et al.
Published: (2023)
by: Mathieu, Timothée, et al.
Published: (2023)
Evaluating Interpretable Reinforcement Learning by Distilling Policies into Programs
by: Kohler, Hector, et al.
Published: (2025)
by: Kohler, Hector, et al.
Published: (2025)
Interpolation pour l'augmentation de donnees : Application à la gestion des adventices de la canne a sucre a la Reunion
by: Ferber, Frederick Fabre, et al.
Published: (2025)
by: Ferber, Frederick Fabre, et al.
Published: (2025)
Abstraction for Offline Goal-Conditioned Reinforcement Learning
by: Wibault, Clarisse, et al.
Published: (2026)
by: Wibault, Clarisse, et al.
Published: (2026)
Latent Representation Alignment for Offline Goal-Conditioned Reinforcement Learning
by: Kang, Hyungkyu, et al.
Published: (2026)
by: Kang, Hyungkyu, et al.
Published: (2026)
QuasiNav: Asymmetric Cost-Aware Navigation Planning with Constrained Quasimetric Reinforcement Learning
by: Hossain, Jumman, et al.
Published: (2024)
by: Hossain, Jumman, et al.
Published: (2024)
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
by: Huang, Xingshuai, et al.
Published: (2024)
by: Huang, Xingshuai, et al.
Published: (2024)
Power Mean Estimation in Stochastic Monte-Carlo Tree_Search
by: Dam, Tuan, et al.
Published: (2024)
by: Dam, Tuan, et al.
Published: (2024)
Physics-informed Value Learner for Offline Goal-Conditioned Reinforcement Learning
by: Giammarino, Vittorio, et al.
Published: (2025)
by: Giammarino, Vittorio, et al.
Published: (2025)
Compositional Transduction with Latent Analogies for Offline Goal-Conditioned Reinforcement Learning
by: Kim, Junseok, et al.
Published: (2026)
by: Kim, Junseok, et al.
Published: (2026)
GOPlan: Goal-conditioned Offline Reinforcement Learning by Planning with Learned Models
by: Wang, Mianchu, et al.
Published: (2023)
by: Wang, Mianchu, et al.
Published: (2023)
SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
Option-aware Temporally Abstracted Value for Offline Goal-Conditioned Reinforcement Learning
by: Ahn, Hongjoon, et al.
Published: (2025)
by: Ahn, Hongjoon, et al.
Published: (2025)
Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning
by: Venugopal, Aravind, et al.
Published: (2026)
by: Venugopal, Aravind, et al.
Published: (2026)
Adaptive Coarse-to-Fine Subgoal Refinement for Long-Horizon Offline Goal-Conditioned Reinforcement Learning
by: Ke, Kaiqiang, et al.
Published: (2026)
by: Ke, Kaiqiang, et al.
Published: (2026)
Proposing Hierarchical Goal-Conditioned Policy Planning in Multi-Goal Reinforcement Learning
by: Rens, Gavin B.
Published: (2025)
by: Rens, Gavin B.
Published: (2025)
Test-time Offline Reinforcement Learning on Goal-related Experience
by: Bagatella, Marco, et al.
Published: (2025)
by: Bagatella, Marco, et al.
Published: (2025)
Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy
by: Cao, Chenyang, et al.
Published: (2024)
by: Cao, Chenyang, et al.
Published: (2024)
Goal-conditioned Offline Reinforcement Learning through State Space Partitioning
by: Wang, Mianchu, et al.
Published: (2023)
by: Wang, Mianchu, et al.
Published: (2023)
OGBench: Benchmarking Offline Goal-Conditioned RL
by: Park, Seohong, et al.
Published: (2024)
by: Park, Seohong, et al.
Published: (2024)
Quasimetric Value Functions with Dense Rewards
by: Valieva, Khadichabonu, et al.
Published: (2024)
by: Valieva, Khadichabonu, et al.
Published: (2024)
Similar Items
-
A Continual Offline Reinforcement Learning Benchmark for Navigation Tasks
by: Kobanda, Anthony, et al.
Published: (2025) -
Hierarchical Subspaces of Policies for Continual Offline Reinforcement Learning
by: Kobanda, Anthony, et al.
Published: (2024) -
How Hard is it to Confuse a World Model?
by: Radji, Waris, et al.
Published: (2025) -
The Confusing Instance Principle for Online Linear Quadratic Control
by: Radji, Waris, et al.
Published: (2025) -
Intrinsic-Energy Joint Embedding Predictive Architectures Induce Quasimetric Spaces
by: Kobanda, Anthony, et al.
Published: (2026)