Saved in:
| Main Authors: | Sanati, Ben, Lee, Thomas L., McInroe, Trevor, Scannell, Aidan, Malkin, Nikolay, Abel, David, Storkey, Amos |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2511.04666 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Offline Reinforcement Learning: First Imitate, then Improve
by: Jelley, Adam, et al.
Published: (2024)
by: Jelley, Adam, et al.
Published: (2024)
Planning to Go Out-of-Distribution in Offline-to-Online Reinforcement Learning
by: McInroe, Trevor, et al.
Published: (2023)
by: McInroe, Trevor, et al.
Published: (2023)
Object-Centric World Models from Few-Shot Annotations for Sample-Efficient Reinforcement Learning
by: Zhang, Weipu, et al.
Published: (2025)
by: Zhang, Weipu, et al.
Published: (2025)
PixelBrax: Learning Continuous Control from Pixels End-to-End on the GPU
by: McInroe, Trevor, et al.
Published: (2025)
by: McInroe, Trevor, et al.
Published: (2025)
Multi-Horizon Representations with Hierarchical Forward Models for Reinforcement Learning
by: McInroe, Trevor, et al.
Published: (2022)
by: McInroe, Trevor, et al.
Published: (2022)
Enhancing Tactile-based Reinforcement Learning for Robotic Control
by: Miller, Elle, et al.
Published: (2025)
by: Miller, Elle, et al.
Published: (2025)
Terra Nova: A Comprehensive Challenge Environment for Intelligent Agents
by: McInroe, Trevor
Published: (2025)
by: McInroe, Trevor
Published: (2025)
Kalman Linear Attention: Parallel Bayesian Filtering For Efficient Language Modelling and State Tracking
by: Shaj, Vaisakh, et al.
Published: (2026)
by: Shaj, Vaisakh, et al.
Published: (2026)
roto 2.0: The Robot Tactile Olympiad
by: Miller, Elle, et al.
Published: (2026)
by: Miller, Elle, et al.
Published: (2026)
Approximate Bayesian Class-Conditional Models under Continuous Representation Shift
by: Lee, Thomas L., et al.
Published: (2023)
by: Lee, Thomas L., et al.
Published: (2023)
Chunking: Continual Learning is not just about Distribution Shift
by: Lee, Thomas L., et al.
Published: (2023)
by: Lee, Thomas L., et al.
Published: (2023)
LLM-Personalize: Aligning LLM Planners with Human Preferences via Reinforced Self-Training for Housekeeping Robots
by: Han, Dongge, et al.
Published: (2024)
by: Han, Dongge, et al.
Published: (2024)
Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning
by: Garcin, Samuel, et al.
Published: (2025)
by: Garcin, Samuel, et al.
Published: (2025)
Adapting Time Series Foundation Models through Data Mixtures
by: Lee, Thomas L., et al.
Published: (2026)
by: Lee, Thomas L., et al.
Published: (2026)
Generative World Modelling for Humanoids: 1X World Model Challenge Technical Report
by: Mereu, Riccardo, et al.
Published: (2025)
by: Mereu, Riccardo, et al.
Published: (2025)
Label Noise: Correcting the Forward-Correction
by: Toner, William, et al.
Published: (2023)
by: Toner, William, et al.
Published: (2023)
Noisy Early Stopping for Noisy Labels
by: Toner, William, et al.
Published: (2024)
by: Toner, William, et al.
Published: (2024)
Probing Dec-POMDP Reasoning in Cooperative MARL
by: Tessera, Kale-ab, et al.
Published: (2026)
by: Tessera, Kale-ab, et al.
Published: (2026)
CODA: Coordination via On-Policy Diffusion for Multi-Agent Offline Reinforcement Learning
by: Hedman, Marcel, et al.
Published: (2026)
by: Hedman, Marcel, et al.
Published: (2026)
Adversarial robustness of VAEs through the lens of local geometry
by: Khan, Asif, et al.
Published: (2022)
by: Khan, Asif, et al.
Published: (2022)
Rationality Measurement and Theory for Reinforcement Learning Agents
by: Qian, Kejiang, et al.
Published: (2026)
by: Qian, Kejiang, et al.
Published: (2026)
Remembering the Markov Property in Cooperative MARL
by: Tessera, Kale-ab Abebe, et al.
Published: (2025)
by: Tessera, Kale-ab Abebe, et al.
Published: (2025)
Entropy Regularized Task Representation Learning for Offline Meta-Reinforcement Learning
by: Nakhaei, Mohammadreza, et al.
Published: (2024)
by: Nakhaei, Mohammadreza, et al.
Published: (2024)
Residual Learning and Context Encoding for Adaptive Offline-to-Online Reinforcement Learning
by: Nakhaei, Mohammadreza, et al.
Published: (2024)
by: Nakhaei, Mohammadreza, et al.
Published: (2024)
Hyperparameter Selection in Continual Learning
by: Lee, Thomas L., et al.
Published: (2024)
by: Lee, Thomas L., et al.
Published: (2024)
Assistax: A Hardware-Accelerated Reinforcement Learning Benchmark for Assistive Robotics
by: Hinckeldey, Leonard, et al.
Published: (2025)
by: Hinckeldey, Leonard, et al.
Published: (2025)
Data-to-Energy Stochastic Dynamics
by: Tamogashev, Kirill, et al.
Published: (2025)
by: Tamogashev, Kirill, et al.
Published: (2025)
Contextual Latent World Models for Offline Meta Reinforcement Learning
by: Nakheai, Mohammadreza, et al.
Published: (2026)
by: Nakheai, Mohammadreza, et al.
Published: (2026)
Meta-Adaptive Prompt Distillation for Few-Shot Visual Question Answering
by: Gupta, Akash, et al.
Published: (2025)
by: Gupta, Akash, et al.
Published: (2025)
On Designing Diffusion Autoencoders for Efficient Generation and Representation Learning
by: Proszewska, Magdalena, et al.
Published: (2025)
by: Proszewska, Magdalena, et al.
Published: (2025)
Few-Shot Learning with Class Imbalance
by: Ochal, Mateusz, et al.
Published: (2021)
by: Ochal, Mateusz, et al.
Published: (2021)
iQRL -- Implicitly Quantized Representations for Sample-efficient Reinforcement Learning
by: Scannell, Aidan, et al.
Published: (2024)
by: Scannell, Aidan, et al.
Published: (2024)
Function-space Parameterization of Neural Networks for Sequential Learning
by: Scannell, Aidan, et al.
Published: (2024)
by: Scannell, Aidan, et al.
Published: (2024)
On Harnessing Idle Compute at the Edge for Foundation Model Training
by: Xue, Leyang, et al.
Published: (2025)
by: Xue, Leyang, et al.
Published: (2025)
Signature-Kernel Based Evaluation Metrics for Robust Probabilistic and Tail-Event Forecasting
by: Redhead, Benjamin R., et al.
Published: (2026)
by: Redhead, Benjamin R., et al.
Published: (2026)
On Generalization for Generative Flow Networks
by: Krichel, Anas, et al.
Published: (2024)
by: Krichel, Anas, et al.
Published: (2024)
Aligning Agents like Large Language Models
by: Jelley, Adam, et al.
Published: (2024)
by: Jelley, Adam, et al.
Published: (2024)
HyperMARL: Adaptive Hypernetworks for Multi-Agent RL
by: Tessera, Kale-ab Abebe, et al.
Published: (2024)
by: Tessera, Kale-ab Abebe, et al.
Published: (2024)
In-Context Parametric Inference: Point or Distribution Estimators?
by: Mittal, Sarthak, et al.
Published: (2025)
by: Mittal, Sarthak, et al.
Published: (2025)
Discrete Codebook World Models for Continuous Control
by: Scannell, Aidan, et al.
Published: (2025)
by: Scannell, Aidan, et al.
Published: (2025)
Similar Items
-
Efficient Offline Reinforcement Learning: First Imitate, then Improve
by: Jelley, Adam, et al.
Published: (2024) -
Planning to Go Out-of-Distribution in Offline-to-Online Reinforcement Learning
by: McInroe, Trevor, et al.
Published: (2023) -
Object-Centric World Models from Few-Shot Annotations for Sample-Efficient Reinforcement Learning
by: Zhang, Weipu, et al.
Published: (2025) -
PixelBrax: Learning Continuous Control from Pixels End-to-End on the GPU
by: McInroe, Trevor, et al.
Published: (2025) -
Multi-Horizon Representations with Hierarchical Forward Models for Reinforcement Learning
by: McInroe, Trevor, et al.
Published: (2022)