Q-Learning under Finite Model Uncertainty
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sester, Julian, Decker, Cécile |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Robust $Q$-learning Algorithm for Markov Decision Processes under Wasserstein Uncertainty
von: Neufeld, Ariel, et al.
Veröffentlicht: (2022)
von: Neufeld, Ariel, et al.
Veröffentlicht: (2022)
Probabilistic Geometric Alignment via Bayesian Latent Transport for Domain-Adaptive Foundation Models
von: Aueawatthanaphisut, Aueaphum, et al.
Veröffentlicht: (2026)
von: Aueawatthanaphisut, Aueaphum, et al.
Veröffentlicht: (2026)
Lagrangian Index Policy for Restless Bandits with Average Reward
von: Avrachenkov, Konstantin, et al.
Veröffentlicht: (2024)
von: Avrachenkov, Konstantin, et al.
Veröffentlicht: (2024)
On the consistent reasoning paradox of intelligence and optimal trust in AI: The power of 'I don't know'
von: Bastounis, Alexander, et al.
Veröffentlicht: (2024)
von: Bastounis, Alexander, et al.
Veröffentlicht: (2024)
Neural Brownian Motion
von: Qi, Qian
Veröffentlicht: (2025)
von: Qi, Qian
Veröffentlicht: (2025)
Feature-aligned N-BEATS with Sinkhorn divergence
von: Lee, Joonhun, et al.
Veröffentlicht: (2023)
von: Lee, Joonhun, et al.
Veröffentlicht: (2023)
Efficient Risk-sensitive Planning via Entropic Risk Measures
von: Marthe, Alexandre, et al.
Veröffentlicht: (2025)
von: Marthe, Alexandre, et al.
Veröffentlicht: (2025)
Bounding the Difference between the Values of Robust and Non-Robust Markov Decision Problems
von: Neufeld, Ariel, et al.
Veröffentlicht: (2023)
von: Neufeld, Ariel, et al.
Veröffentlicht: (2023)
Non-concave stochastic optimal control in finite discrete time under model uncertainty
von: Neufeld, Ariel, et al.
Veröffentlicht: (2024)
von: Neufeld, Ariel, et al.
Veröffentlicht: (2024)
Distributionally Robust Deep Q-Learning
von: Lu, Chung I, et al.
Veröffentlicht: (2025)
von: Lu, Chung I, et al.
Veröffentlicht: (2025)
DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty
von: Cui, Mingxuan, et al.
Veröffentlicht: (2025)
von: Cui, Mingxuan, et al.
Veröffentlicht: (2025)
Deep Learning for Sequential Decision Making under Uncertainty: Foundations, Frameworks, and Frontiers
von: Buyuktahtakin, I. Esra
Veröffentlicht: (2026)
von: Buyuktahtakin, I. Esra
Veröffentlicht: (2026)
Adaptively Robust LLM Inference Optimization under Prediction Uncertainty
von: Chen, Zixi, et al.
Veröffentlicht: (2025)
von: Chen, Zixi, et al.
Veröffentlicht: (2025)
AI2STOW: End-to-End Deep Reinforcement Learning to Construct Master Stowage Plans under Demand Uncertainty
von: Van Twiller, Jaike, et al.
Veröffentlicht: (2025)
von: Van Twiller, Jaike, et al.
Veröffentlicht: (2025)
Robust Reinforcement Learning in Finance: Modeling Market Impact with Elliptic Uncertainty Sets
von: Ma, Shaocong, et al.
Veröffentlicht: (2025)
von: Ma, Shaocong, et al.
Veröffentlicht: (2025)
Wasserstein Convergence of Score-based Generative Models under Semiconvexity and Discontinuous Gradients
von: Bruno, Stefano, et al.
Veröffentlicht: (2025)
von: Bruno, Stefano, et al.
Veröffentlicht: (2025)
Diffusion-Based Stochastic Operator Networks for Uncertainty Quantification in Stochastic Partial Differential Equations
von: Huynh, Phuoc-Toan, et al.
Veröffentlicht: (2026)
von: Huynh, Phuoc-Toan, et al.
Veröffentlicht: (2026)
Pairwise independent correlation gap
von: Ramachandra, Arjun, et al.
Veröffentlicht: (2022)
von: Ramachandra, Arjun, et al.
Veröffentlicht: (2022)
Universal Approximation Theorem for Deep Q-Learning via FBSDE System
von: Qi, Qian
Veröffentlicht: (2025)
von: Qi, Qian
Veröffentlicht: (2025)
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
von: Zhu, Feng, et al.
Veröffentlicht: (2025)
von: Zhu, Feng, et al.
Veröffentlicht: (2025)
An Improved Finite-time Analysis of Temporal Difference Learning with Deep Neural Networks
von: Ke, Zhifa, et al.
Veröffentlicht: (2024)
von: Ke, Zhifa, et al.
Veröffentlicht: (2024)
Algorithmic Stability of Stochastic Gradient Descent with Momentum under Heavy-Tailed Noise
von: Dang, Thanh, et al.
Veröffentlicht: (2025)
von: Dang, Thanh, et al.
Veröffentlicht: (2025)
Robust Control with Gradient Uncertainty
von: Qi, Qian
Veröffentlicht: (2025)
von: Qi, Qian
Veröffentlicht: (2025)
The Gittins Index: A Design Principle for Decision-Making Under Uncertainty
von: Scully, Ziv, et al.
Veröffentlicht: (2025)
von: Scully, Ziv, et al.
Veröffentlicht: (2025)
Finite-Time Analysis of Gradient Descent for Shallow Transformers
von: Arda, Enes, et al.
Veröffentlicht: (2026)
von: Arda, Enes, et al.
Veröffentlicht: (2026)
Provable Acceleration for Diffusion Models under Minimal Assumptions
von: Li, Gen, et al.
Veröffentlicht: (2024)
von: Li, Gen, et al.
Veröffentlicht: (2024)
Value Mirror Descent for Reinforcement Learning
von: Jia, Zhichao, et al.
Veröffentlicht: (2026)
von: Jia, Zhichao, et al.
Veröffentlicht: (2026)
Reinforcement Learning with Random Time Horizons
von: Borrell, Enric Ribera, et al.
Veröffentlicht: (2025)
von: Borrell, Enric Ribera, et al.
Veröffentlicht: (2025)
Reinforcement Learning under Latent Dynamics: Toward Statistical and Algorithmic Modularity
von: Amortila, Philip, et al.
Veröffentlicht: (2024)
von: Amortila, Philip, et al.
Veröffentlicht: (2024)
A Generalization Result for Convergence in Learning-to-Optimize
von: Sucker, Michael, et al.
Veröffentlicht: (2024)
von: Sucker, Michael, et al.
Veröffentlicht: (2024)
Multi-Year Maintenance Planning for Large-Scale Infrastructure Systems: A Novel Network Deep Q-Learning Approach
von: Fard, Amir, et al.
Veröffentlicht: (2025)
von: Fard, Amir, et al.
Veröffentlicht: (2025)
Learning-Based Pricing and Matching for Two-Sided Queues
von: Yang, Zixian, et al.
Veröffentlicht: (2024)
von: Yang, Zixian, et al.
Veröffentlicht: (2024)
Asymptotic and Finite Sample Analysis of Nonexpansive Stochastic Approximations with Markovian Noise
von: Blaser, Ethan, et al.
Veröffentlicht: (2024)
von: Blaser, Ethan, et al.
Veröffentlicht: (2024)
Non-Smooth Weakly-Convex Finite-sum Coupled Compositional Optimization
von: Hu, Quanqi, et al.
Veröffentlicht: (2023)
von: Hu, Quanqi, et al.
Veröffentlicht: (2023)
Algebraic Reduction of Hidden Markov Models
von: Grigoletto, Tommaso, et al.
Veröffentlicht: (2022)
von: Grigoletto, Tommaso, et al.
Veröffentlicht: (2022)
Admission Control of Quasi-Reversible Queueing Systems: Optimization and Reinforcement Learning
von: Comte, Céline, et al.
Veröffentlicht: (2025)
von: Comte, Céline, et al.
Veröffentlicht: (2025)
Online Learning and Optimization for Queues with Unknown Demand Curve and Service Distribution
von: Chen, Xinyun, et al.
Veröffentlicht: (2023)
von: Chen, Xinyun, et al.
Veröffentlicht: (2023)
Wasserstein Formulation of Reinforcement Learning. An Optimal Transport Perspective on Policy Optimization
von: Dus, Mathias
Veröffentlicht: (2026)
von: Dus, Mathias
Veröffentlicht: (2026)
Fitted Q-Iteration via Max-Plus-Linear Approximation
von: Liu, Y., et al.
Veröffentlicht: (2024)
von: Liu, Y., et al.
Veröffentlicht: (2024)
Q3R: Quadratic Reweighted Rank Regularizer for Effective Low-Rank Training
von: Ghosh, Ipsita, et al.
Veröffentlicht: (2025)
von: Ghosh, Ipsita, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Robust $Q$-learning Algorithm for Markov Decision Processes under Wasserstein Uncertainty
von: Neufeld, Ariel, et al.
Veröffentlicht: (2022) -
Probabilistic Geometric Alignment via Bayesian Latent Transport for Domain-Adaptive Foundation Models
von: Aueawatthanaphisut, Aueaphum, et al.
Veröffentlicht: (2026) -
Lagrangian Index Policy for Restless Bandits with Average Reward
von: Avrachenkov, Konstantin, et al.
Veröffentlicht: (2024) -
On the consistent reasoning paradox of intelligence and optimal trust in AI: The power of 'I don't know'
von: Bastounis, Alexander, et al.
Veröffentlicht: (2024) -
Neural Brownian Motion
von: Qi, Qian
Veröffentlicht: (2025)