Reinforcement learning with non-ergodic reward increments: robustness via ergodicity transformations
Fuente:
arXiv
Saved in:
| Main Authors: | Baumann, Dominik, Noorani, Erfaun, Price, James, Peters, Ole, Connaughton, Colm, Schön, Thomas B. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Ergodicity in reinforcement learning
by: Baumann, Dominik, et al.
Published: (2026)
by: Baumann, Dominik, et al.
Published: (2026)
Risk-Sensitive Reinforcement Learning with Exponential Criteria
by: Noorani, Erfaun, et al.
Published: (2022)
by: Noorani, Erfaun, et al.
Published: (2022)
Safe reinforcement learning in uncertain contexts
by: Baumann, Dominik, et al.
Published: (2024)
by: Baumann, Dominik, et al.
Published: (2024)
From Abstraction to Reality: DARPA's Vision for Robust Sim-to-Real Autonomy
by: Noorani, Erfaun, et al.
Published: (2025)
by: Noorani, Erfaun, et al.
Published: (2025)
Distributed Risk-Sensitive Safety Filters for Uncertain Discrete-Time Systems
by: Lederer, Armin, et al.
Published: (2025)
by: Lederer, Armin, et al.
Published: (2025)
Safe learning-based control via function-based uncertainty quantification
by: Tokmak, Abdullah, et al.
Published: (2026)
by: Tokmak, Abdullah, et al.
Published: (2026)
Safe Bayesian optimization across noise models via scenario programming
by: Tokmak, Abdullah, et al.
Published: (2025)
by: Tokmak, Abdullah, et al.
Published: (2025)
Towards Efficient Risk-Sensitive Policy Gradient: An Iteration Complexity Analysis
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
Counterfactual Explanations for Model Ensembles Using Entropic Risk Measures
by: Noorani, Erfaun, et al.
Published: (2025)
by: Noorani, Erfaun, et al.
Published: (2025)
From abstraction to reality: DARPA's vision for robust sim‐to‐real autonomy
by: Erfaun Noorani, et al.
Published: (2025)
by: Erfaun Noorani, et al.
Published: (2025)
Towards safe control parameter tuning in distributed multi-agent systems
by: Tokmak, Abdullah, et al.
Published: (2025)
by: Tokmak, Abdullah, et al.
Published: (2025)
PACSBO: Probably approximately correct safe Bayesian optimization
by: Tokmak, Abdullah, et al.
Published: (2024)
by: Tokmak, Abdullah, et al.
Published: (2024)
Robust Counterfactual Explanations for Neural Networks With Probabilistic Guarantees
by: Hamman, Faisal, et al.
Published: (2023)
by: Hamman, Faisal, et al.
Published: (2023)
Safe exploration in reproducing kernel Hilbert spaces
by: Tokmak, Abdullah, et al.
Published: (2025)
by: Tokmak, Abdullah, et al.
Published: (2025)
A non-ergodic framework for understanding emergent capabilities in Large Language Models
by: Marín, Javier
Published: (2025)
by: Marín, Javier
Published: (2025)
Geometric ergodicity of SGLD via reflection coupling
by: Li, Lei, et al.
Published: (2023)
by: Li, Lei, et al.
Published: (2023)
On the strong stability of ergodic iterations
by: Györfi, László, et al.
Published: (2023)
by: Györfi, László, et al.
Published: (2023)
Deep neural networks from the perspective of ergodic theory
by: Zhang, Fan
Published: (2023)
by: Zhang, Fan
Published: (2023)
Beyond expected value: geometric mean optimization for long-term policy performance in reinforcement learning
by: Sheng, Xinyi, et al.
Published: (2025)
by: Sheng, Xinyi, et al.
Published: (2025)
Advancing Robustness in Deep Reinforcement Learning with an Ensemble Defense Approach
by: Mohan, Adithya, et al.
Published: (2025)
by: Mohan, Adithya, et al.
Published: (2025)
Optimistic Q-learning for average reward and episodic reinforcement learning
by: Agrawal, Priyank, et al.
Published: (2024)
by: Agrawal, Priyank, et al.
Published: (2024)
Active teacher selection for reward learning
by: Freedman, Rachel, et al.
Published: (2023)
by: Freedman, Rachel, et al.
Published: (2023)
Noise-based reward-modulated learning
by: Fernández, Jesús García, et al.
Published: (2025)
by: Fernández, Jesús García, et al.
Published: (2025)
State space models, emergence, and ergodicity: How many parameters are needed for stable predictions?
by: Ziemann, Ingvar, et al.
Published: (2024)
by: Ziemann, Ingvar, et al.
Published: (2024)
A Rigorous, Tractable Measure of Model Complexity
by: Allerbo, Oskar, et al.
Published: (2026)
by: Allerbo, Oskar, et al.
Published: (2026)
Is Supervised Learning Really That Different from Unsupervised?
by: Allerbo, Oskar, et al.
Published: (2025)
by: Allerbo, Oskar, et al.
Published: (2025)
Priority-Driven Control and Communication in Decentralized Multi-Agent Systems via Reinforcement Learning
by: Guo, Qingyun, et al.
Published: (2026)
by: Guo, Qingyun, et al.
Published: (2026)
Convergence of Adam for Non-convex Objectives: Relaxed Hyperparameters and Non-ergodic Case
by: He, Meixuan, et al.
Published: (2023)
by: He, Meixuan, et al.
Published: (2023)
Correction to "Wasserstein distance estimates for the distributions of numerical approximations to ergodic stochastic differential equations"
by: Paulin, Daniel, et al.
Published: (2024)
by: Paulin, Daniel, et al.
Published: (2024)
The impact of intrinsic rewards on exploration in Reinforcement Learning
by: Kayal, Aya, et al.
Published: (2025)
by: Kayal, Aya, et al.
Published: (2025)
Episodic Reinforcement Learning with Expanded State-reward Space
by: Liang, Dayang, et al.
Published: (2024)
by: Liang, Dayang, et al.
Published: (2024)
Multi-task learning via robust regularized clustering with non-convex group penalties
by: Okazaki, Akira, et al.
Published: (2024)
by: Okazaki, Akira, et al.
Published: (2024)
Operating critical machine learning models in resource constrained regimes
by: Selvan, Raghavendra, et al.
Published: (2023)
by: Selvan, Raghavendra, et al.
Published: (2023)
Entropy-regularized Diffusion Policy with Q-Ensembles for Offline Reinforcement Learning
by: Zhang, Ruoqi, et al.
Published: (2024)
by: Zhang, Ruoqi, et al.
Published: (2024)
Unsupervised dynamic modeling of medical image transformation
by: Gunnarsson, Niklas, et al.
Published: (2021)
by: Gunnarsson, Niklas, et al.
Published: (2021)
Revisiting Value Iteration: Unified Analysis of Discounted and Average-Reward Cases
by: Mustafin, Arsenii, et al.
Published: (2025)
by: Mustafin, Arsenii, et al.
Published: (2025)
Transfer Learning in Latent Contextual Bandits with Covariate Shift Through Causal Transportability
by: Deng, Mingwei, et al.
Published: (2025)
by: Deng, Mingwei, et al.
Published: (2025)
Multi-Round Human-AI Collaboration with User-Specified Requirements
by: Noorani, Sima, et al.
Published: (2026)
by: Noorani, Sima, et al.
Published: (2026)
Lecture notes on ergodic transformations
by: Ryzhikov, Valery V.
Published: (2024)
by: Ryzhikov, Valery V.
Published: (2024)
Reproducibility study on how to find Spurious Correlations, Shortcut Learning, Clever Hans or Group-Distributional non-robustness and how to fix them
by: Delzer, Ole, et al.
Published: (2026)
by: Delzer, Ole, et al.
Published: (2026)
Similar Items
-
Ergodicity in reinforcement learning
by: Baumann, Dominik, et al.
Published: (2026) -
Risk-Sensitive Reinforcement Learning with Exponential Criteria
by: Noorani, Erfaun, et al.
Published: (2022) -
Safe reinforcement learning in uncertain contexts
by: Baumann, Dominik, et al.
Published: (2024) -
From Abstraction to Reality: DARPA's Vision for Robust Sim-to-Real Autonomy
by: Noorani, Erfaun, et al.
Published: (2025) -
Distributed Risk-Sensitive Safety Filters for Uncertain Discrete-Time Systems
by: Lederer, Armin, et al.
Published: (2025)