Saved in:
| Main Authors: | Frost, Thomas, Vaidya, Hrisheekesh, Harris, Steve |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2602.06603 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robust Real-Time Mortality Prediction in the Intensive Care Unit using Temporal Difference Learning
by: Frost, Thomas, et al.
Published: (2024)
by: Frost, Thomas, et al.
Published: (2024)
The challenge of hidden gifts in multi-agent reinforcement learning
by: Malenfant, Dane, et al.
Published: (2025)
by: Malenfant, Dane, et al.
Published: (2025)
Variational predictive resampling
by: Battaglia, Laura, et al.
Published: (2026)
by: Battaglia, Laura, et al.
Published: (2026)
Ergodicity in reinforcement learning
by: Baumann, Dominik, et al.
Published: (2026)
by: Baumann, Dominik, et al.
Published: (2026)
Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning
by: Kobayashi, Seijin, et al.
Published: (2025)
by: Kobayashi, Seijin, et al.
Published: (2025)
Risk-sensitive reinforcement learning using expectiles, shortfall risk and optimized certainty equivalent risk
by: Gupte, Sumedh, et al.
Published: (2026)
by: Gupte, Sumedh, et al.
Published: (2026)
Safe reinforcement learning in uncertain contexts
by: Baumann, Dominik, et al.
Published: (2024)
by: Baumann, Dominik, et al.
Published: (2024)
Universal hidden monotonic trend estimation with contrastive learning
by: Pineau, Edouard, et al.
Published: (2022)
by: Pineau, Edouard, et al.
Published: (2022)
Maximum diffusion reinforcement learning
by: Berrueta, Thomas A., et al.
Published: (2023)
by: Berrueta, Thomas A., et al.
Published: (2023)
mldr.resampling: Efficient Reference Implementations of Multilabel Resampling Algorithms
by: Rivera, Antonio J., et al.
Published: (2023)
by: Rivera, Antonio J., et al.
Published: (2023)
TemporalPaD: a reinforcement-learning framework for temporal feature representation and dimension reduction
by: Mu, Xuechen, et al.
Published: (2024)
by: Mu, Xuechen, et al.
Published: (2024)
Catastrophic-risk-aware reinforcement learning with extreme-value-theory-based policy gradients
by: Davar, Parisa, et al.
Published: (2024)
by: Davar, Parisa, et al.
Published: (2024)
When resampling/reweighting improves feature learning in imbalanced classification?: A toy-model study
by: Obuchi, Tomoyuki, et al.
Published: (2024)
by: Obuchi, Tomoyuki, et al.
Published: (2024)
Residual resampling-based physics-informed neural network for neutron diffusion equations
by: Zhang, Heng, et al.
Published: (2024)
by: Zhang, Heng, et al.
Published: (2024)
MPCritic: A plug-and-play MPC architecture for reinforcement learning
by: Lawrence, Nathan P., et al.
Published: (2025)
by: Lawrence, Nathan P., et al.
Published: (2025)
Normalization and effective learning rates in reinforcement learning
by: Lyle, Clare, et al.
Published: (2024)
by: Lyle, Clare, et al.
Published: (2024)
An introduction to reinforcement learning for neuroscience
by: Jensen, Kristopher T.
Published: (2023)
by: Jensen, Kristopher T.
Published: (2023)
Optimistic Q-learning for average reward and episodic reinforcement learning
by: Agrawal, Priyank, et al.
Published: (2024)
by: Agrawal, Priyank, et al.
Published: (2024)
Variational Autoencoders for exteroceptive perception in reinforcement learning-based collision avoidance
by: Larsen, Thomas Nakken, et al.
Published: (2024)
by: Larsen, Thomas Nakken, et al.
Published: (2024)
Simple and near-optimal algorithms for hidden stratification and multi-group learning
by: Tosh, Christopher, et al.
Published: (2021)
by: Tosh, Christopher, et al.
Published: (2021)
A method of supervised learning from conflicting data with hidden contexts
by: Zhang, Tianren, et al.
Published: (2021)
by: Zhang, Tianren, et al.
Published: (2021)
Shallow diffusion networks provably learn hidden low-dimensional structure
by: Boffi, Nicholas M., et al.
Published: (2024)
by: Boffi, Nicholas M., et al.
Published: (2024)
Generalized Bayesian deep reinforcement learning
by: Roy, Shreya Sinha, et al.
Published: (2024)
by: Roy, Shreya Sinha, et al.
Published: (2024)
Universal rates of ERM for agnostic learning
by: Hanneke, Steve, et al.
Published: (2025)
by: Hanneke, Steve, et al.
Published: (2025)
Deep autoregressive density nets vs neural ensembles for model-based offline reinforcement learning
by: Benechehab, Abdelhakim, et al.
Published: (2024)
by: Benechehab, Abdelhakim, et al.
Published: (2024)
Curriculum reinforcement learning with measurable task representation learning
by: Wen, Yongyan, et al.
Published: (2026)
by: Wen, Yongyan, et al.
Published: (2026)
What is the objective of reasoning with reinforcement learning?
by: Davis, Damek, et al.
Published: (2025)
by: Davis, Damek, et al.
Published: (2025)
Meta-reinforcement learning with minimum attention
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
Controllability in preference-conditioned multi-objective reinforcement learning
by: Molins, Pau de las Heras, et al.
Published: (2026)
by: Molins, Pau de las Heras, et al.
Published: (2026)
Dragonfly: a modular deep reinforcement learning library
by: Viquerat, Jonathan, et al.
Published: (2025)
by: Viquerat, Jonathan, et al.
Published: (2025)
Expert or not? assessing data quality in offline reinforcement learning
by: Asadulaev, Arip, et al.
Published: (2025)
by: Asadulaev, Arip, et al.
Published: (2025)
Adaptive sampling using variational autoencoder and reinforcement learning
by: Rasheed, Adil, et al.
Published: (2025)
by: Rasheed, Adil, et al.
Published: (2025)
Position: Lifetime tuning is incompatible with continual reinforcement learning
by: Mesbahi, Golnaz, et al.
Published: (2024)
by: Mesbahi, Golnaz, et al.
Published: (2024)
Quantum reinforcement learning in continuous action space
by: Wu, Shaojun, et al.
Published: (2020)
by: Wu, Shaojun, et al.
Published: (2020)
Data-assimilated model-informed reinforcement learning
by: Ozan, Defne E., et al.
Published: (2025)
by: Ozan, Defne E., et al.
Published: (2025)
Current applications and potential future directions of reinforcement learning-based Digital Twins in agriculture
by: Goldenits, Georg, et al.
Published: (2024)
by: Goldenits, Georg, et al.
Published: (2024)
Fundamentals of quantum Boltzmann machine learning with visible and hidden units
by: Wilde, Mark M.
Published: (2025)
by: Wilde, Mark M.
Published: (2025)
Analysing zero-shot temporal relation extraction on clinical notes using temporal consistency
by: Kougia, Vasiliki, et al.
Published: (2024)
by: Kougia, Vasiliki, et al.
Published: (2024)
A perspective on fluid mechanical environments for challenges in reinforcement learning
by: Mishra, Shruti, et al.
Published: (2026)
by: Mishra, Shruti, et al.
Published: (2026)
A deep reinforcement learning platform for antibiotic discovery
by: Cao, Hanqun, et al.
Published: (2025)
by: Cao, Hanqun, et al.
Published: (2025)
Similar Items
-
Robust Real-Time Mortality Prediction in the Intensive Care Unit using Temporal Difference Learning
by: Frost, Thomas, et al.
Published: (2024) -
The challenge of hidden gifts in multi-agent reinforcement learning
by: Malenfant, Dane, et al.
Published: (2025) -
Variational predictive resampling
by: Battaglia, Laura, et al.
Published: (2026) -
Ergodicity in reinforcement learning
by: Baumann, Dominik, et al.
Published: (2026) -
Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning
by: Kobayashi, Seijin, et al.
Published: (2025)