Imagine Beyond! Distributionally Robust Auto-Encoding for State Space Coverage in Online Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Castanet, Nicolas, Sigaud, Olivier, Lamprier, Sylvain |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
by: Gaven, Loris, et al.
Published: (2024)
by: Gaven, Loris, et al.
Published: (2024)
Offline Reinforcement Learning of High-Quality Behaviors Under Robust Style Alignment
by: Petitbois, Mathieu, et al.
Published: (2026)
by: Petitbois, Mathieu, et al.
Published: (2026)
Physics-Informed Model and Hybrid Planning for Efficient Dyna-Style Reinforcement Learning
by: Asri, Zakariae El, et al.
Published: (2024)
by: Asri, Zakariae El, et al.
Published: (2024)
Grounding Large Language Models in Interactive Environments with Online Reinforcement Learning
by: Carta, Thomas, et al.
Published: (2023)
by: Carta, Thomas, et al.
Published: (2023)
Robust Deep Reinforcement Learning Through Adversarial Attacks and Training : A Survey
by: Schott, Lucas, et al.
Published: (2024)
by: Schott, Lucas, et al.
Published: (2024)
Autotelic Agents with Intrinsically Motivated Goal-Conditioned Reinforcement Learning: a Short Survey
by: Colas, Cédric, et al.
Published: (2020)
by: Colas, Cédric, et al.
Published: (2020)
Offline Learning of Controllable Diverse Behaviors
by: Petitbois, Mathieu, et al.
Published: (2025)
by: Petitbois, Mathieu, et al.
Published: (2025)
Distributionally Robust Model-based Reinforcement Learning with Large State Spaces
by: Ramesh, Shyam Sundhar, et al.
Published: (2023)
by: Ramesh, Shyam Sundhar, et al.
Published: (2023)
Navigation with QPHIL: Quantizing Planner for Hierarchical Implicit Q-Learning
by: Canesse, Alexi, et al.
Published: (2024)
by: Canesse, Alexi, et al.
Published: (2024)
A tale of two goals: leveraging sequentiality in multi-goal scenarios
by: Serris, Olivier, et al.
Published: (2025)
by: Serris, Olivier, et al.
Published: (2025)
A Transformer Model for Predicting Chemical Products from Generic SMARTS Templates with Data Augmentation
by: Ozer, Derin, et al.
Published: (2025)
by: Ozer, Derin, et al.
Published: (2025)
Reinforcement Learning for Aligning Large Language Models Agents with Interactive Environments: Quantifying and Mitigating Prompt Overfitting
by: Aissi, Mohamed Salim, et al.
Published: (2024)
by: Aissi, Mohamed Salim, et al.
Published: (2024)
ACT: Agentic Classification Tree
by: Grari, Vincent, et al.
Published: (2025)
by: Grari, Vincent, et al.
Published: (2025)
Imagination-Limited Q-Learning for Offline Reinforcement Learning
by: Liu, Wenhui, et al.
Published: (2025)
by: Liu, Wenhui, et al.
Published: (2025)
Accelerated Online Reinforcement Learning using Auxiliary Start State Distributions
by: Mehra, Aman, et al.
Published: (2025)
by: Mehra, Aman, et al.
Published: (2025)
Agentic Adversarial QA for Improving Domain-Specific LLMs
by: Grari, Vincent, et al.
Published: (2026)
by: Grari, Vincent, et al.
Published: (2026)
Reward-Preserving Attacks For Robust Reinforcement Learning
by: Schott, Lucas, et al.
Published: (2026)
by: Schott, Lucas, et al.
Published: (2026)
Sample Complexity of Distributionally Robust Off-Dynamics Reinforcement Learning with Online Interaction
by: He, Yiting, et al.
Published: (2025)
by: He, Yiting, et al.
Published: (2025)
VIPER: Visual Perception and Explainable Reasoning for Sequential Decision-Making
by: Aissi, Mohamed Salim, et al.
Published: (2025)
by: Aissi, Mohamed Salim, et al.
Published: (2025)
Single-Trajectory Distributionally Robust Reinforcement Learning
by: Liang, Zhipeng, et al.
Published: (2023)
by: Liang, Zhipeng, et al.
Published: (2023)
Structural Deep Encoding for Table Question Answering
by: Mouravieff, Raphaël, et al.
Published: (2025)
by: Mouravieff, Raphaël, et al.
Published: (2025)
Robust Reinforcement Learning Objectives for Sequential Recommender Systems
by: Mozifian, Melissa, et al.
Published: (2023)
by: Mozifian, Melissa, et al.
Published: (2023)
HERAKLES: Hierarchical Skill Compilation for Open-ended LLM Agents
by: Carta, Thomas, et al.
Published: (2025)
by: Carta, Thomas, et al.
Published: (2025)
Reaching Consensus in Cooperative Multi-Agent Reinforcement Learning with Goal Imagination
by: Wang, Liangzhou, et al.
Published: (2024)
by: Wang, Liangzhou, et al.
Published: (2024)
On the Geometry of Reinforcement Learning in Continuous State and Action Spaces
by: Tiwari, Saket, et al.
Published: (2022)
by: Tiwari, Saket, et al.
Published: (2022)
Episodic Reinforcement Learning with Expanded State-reward Space
by: Liang, Dayang, et al.
Published: (2024)
by: Liang, Dayang, et al.
Published: (2024)
Enhancing Robustness in Deep Reinforcement Learning: A Lyapunov Exponent Approach
by: Young, Rory, et al.
Published: (2024)
by: Young, Rory, et al.
Published: (2024)
Dense and Diverse Goal Coverage in Multi Goal Reinforcement Learning
by: Singh, Sagalpreet, et al.
Published: (2025)
by: Singh, Sagalpreet, et al.
Published: (2025)
Geometry of Neural Reinforcement Learning in Continuous State and Action Spaces
by: Tiwari, Saket, et al.
Published: (2025)
by: Tiwari, Saket, et al.
Published: (2025)
Offline Reinforcement Learning in Large State Spaces: Algorithms and Guarantees
by: Jiang, Nan, et al.
Published: (2025)
by: Jiang, Nan, et al.
Published: (2025)
Robust Deep Reinforcement Learning with Adaptive Adversarial Perturbations in Action Space
by: Liu, Qianmei, et al.
Published: (2024)
by: Liu, Qianmei, et al.
Published: (2024)
Locality Sensitive Sparse Encoding for Learning World Models Online
by: Liu, Zichen, et al.
Published: (2024)
by: Liu, Zichen, et al.
Published: (2024)
Black-Box Combinatorial Optimization with Order-Invariant Reinforcement Learning
by: Goudet, Olivier, et al.
Published: (2025)
by: Goudet, Olivier, et al.
Published: (2025)
MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces
by: Gaven, Loris, et al.
Published: (2025)
by: Gaven, Loris, et al.
Published: (2025)
On the Complexity of Offline Reinforcement Learning with $Q^\star$-Approximation and Partial Coverage
by: Liu, Haolin, et al.
Published: (2026)
by: Liu, Haolin, et al.
Published: (2026)
Upper and Lower Bounds for Distributionally Robust Off-Dynamics Reinforcement Learning
by: Liu, Zhishuai, et al.
Published: (2024)
by: Liu, Zhishuai, et al.
Published: (2024)
Near-Optimal Partially Observable Reinforcement Learning with Partial Online State Information
by: Shi, Ming, et al.
Published: (2023)
by: Shi, Ming, et al.
Published: (2023)
Improving Offline-to-Online Reinforcement Learning with Q Conditioned State Entropy Exploration
by: Zhang, Ziqi, et al.
Published: (2023)
by: Zhang, Ziqi, et al.
Published: (2023)
Chimera: State Space Models Beyond Sequences
by: Lahoti, Aakash, et al.
Published: (2025)
by: Lahoti, Aakash, et al.
Published: (2025)
Discretizing Continuous Action Space with Unimodal Probability Distributions for On-Policy Reinforcement Learning
by: Zhu, Yuanyang, et al.
Published: (2024)
by: Zhu, Yuanyang, et al.
Published: (2024)
Similar Items
-
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
by: Gaven, Loris, et al.
Published: (2024) -
Offline Reinforcement Learning of High-Quality Behaviors Under Robust Style Alignment
by: Petitbois, Mathieu, et al.
Published: (2026) -
Physics-Informed Model and Hybrid Planning for Efficient Dyna-Style Reinforcement Learning
by: Asri, Zakariae El, et al.
Published: (2024) -
Grounding Large Language Models in Interactive Environments with Online Reinforcement Learning
by: Carta, Thomas, et al.
Published: (2023) -
Robust Deep Reinforcement Learning Through Adversarial Attacks and Training : A Survey
by: Schott, Lucas, et al.
Published: (2024)