UVIP: Model-Free Approach to Evaluate Reinforcement Learning Algorithms
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Belomestny, Denis, Levin, Ilya, Naumov, Alexey, Samsonov, Sergey |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2021
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
High-Order Error Bounds for Markovian LSA with Richardson-Romberg Extrapolation
von: Levin, Ilya, et al.
Veröffentlicht: (2025)
von: Levin, Ilya, et al.
Veröffentlicht: (2025)
Nonasymptotic Analysis of Stochastic Gradient Descent with the Richardson-Romberg Extrapolation
von: Sheshukova, Marina, et al.
Veröffentlicht: (2024)
von: Sheshukova, Marina, et al.
Veröffentlicht: (2024)
SCAFFLSA: Taming Heterogeneity in Federated Linear Stochastic Approximation and TD Learning
von: Mangold, Paul, et al.
Veröffentlicht: (2024)
von: Mangold, Paul, et al.
Veröffentlicht: (2024)
Statistical analysis of Inverse Entropy-regularized Reinforcement Learning
von: Belomestny, Denis, et al.
Veröffentlicht: (2025)
von: Belomestny, Denis, et al.
Veröffentlicht: (2025)
Improved High-Probability Bounds for the Temporal Difference Learning Algorithm via Exponential Stability
von: Samsonov, Sergey, et al.
Veröffentlicht: (2023)
von: Samsonov, Sergey, et al.
Veröffentlicht: (2023)
Gaussian Approximation and Multiplier Bootstrap for Federated Linear Stochastic Approximation
von: Levin, Ilya, et al.
Veröffentlicht: (2026)
von: Levin, Ilya, et al.
Veröffentlicht: (2026)
First Order Methods with Markovian Noise: from Acceleration to Variational Inequalities
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2023)
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2023)
Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent
von: Sheshukova, Marina, et al.
Veröffentlicht: (2025)
von: Sheshukova, Marina, et al.
Veröffentlicht: (2025)
Statistical inference for Linear Stochastic Approximation with Markovian Noise
von: Samsonov, Sergey, et al.
Veröffentlicht: (2025)
von: Samsonov, Sergey, et al.
Veröffentlicht: (2025)
Gaussian Approximation for Two-Timescale Linear Stochastic Approximation
von: Butyrin, Bogdan, et al.
Veröffentlicht: (2025)
von: Butyrin, Bogdan, et al.
Veröffentlicht: (2025)
Theoretical guarantees for neural control variates in MCMC
von: Belomestny, Denis, et al.
Veröffentlicht: (2023)
von: Belomestny, Denis, et al.
Veröffentlicht: (2023)
Rates of convergence for density estimation with generative adversarial networks
von: Puchkin, Nikita, et al.
Veröffentlicht: (2021)
von: Puchkin, Nikita, et al.
Veröffentlicht: (2021)
Refined Analysis of Federated Averaging and Federated Richardson-Romberg
von: Mangold, Paul, et al.
Veröffentlicht: (2024)
von: Mangold, Paul, et al.
Veröffentlicht: (2024)
Gaussian Approximation and Multiplier Bootstrap for Polyak-Ruppert Averaged Linear Stochastic Approximation with Applications to TD Learning
von: Samsonov, Sergey, et al.
Veröffentlicht: (2024)
von: Samsonov, Sergey, et al.
Veröffentlicht: (2024)
Weighted mesh algorithms for general Markov decision processes: Convergence and tractability
von: Belomestny, Denis, et al.
Veröffentlicht: (2024)
von: Belomestny, Denis, et al.
Veröffentlicht: (2024)
A Hessian-Free Actor-Critic Algorithm for Bi-Level Reinforcement Learning with Applications to LLM Fine-Tuning
von: Zeng, Sihan, et al.
Veröffentlicht: (2026)
von: Zeng, Sihan, et al.
Veröffentlicht: (2026)
MPC-Inspired Reinforcement Learning for Verifiable Model-Free Control
von: Lu, Yiwen, et al.
Veröffentlicht: (2023)
von: Lu, Yiwen, et al.
Veröffentlicht: (2023)
Deep Reinforcement Learning: A Convex Optimization Approach
von: Gattami, Ather
Veröffentlicht: (2024)
von: Gattami, Ather
Veröffentlicht: (2024)
Improved Central Limit Theorem and Bootstrap Approximations for Linear Stochastic Approximation
von: Butyrin, Bogdan, et al.
Veröffentlicht: (2025)
von: Butyrin, Bogdan, et al.
Veröffentlicht: (2025)
A Nearly Optimal and Low-Switching Algorithm for Reinforcement Learning with General Function Approximation
von: Zhao, Heyang, et al.
Veröffentlicht: (2023)
von: Zhao, Heyang, et al.
Veröffentlicht: (2023)
Reinforcement Learning Approaches for the Orienteering Problem with Stochastic and Dynamic Release Dates
von: Li, Yuanyuan, et al.
Veröffentlicht: (2022)
von: Li, Yuanyuan, et al.
Veröffentlicht: (2022)
Operator World Models for Reinforcement Learning
von: Novelli, Pietro, et al.
Veröffentlicht: (2024)
von: Novelli, Pietro, et al.
Veröffentlicht: (2024)
Reinforcement Learning for Multi-Truck Vehicle Routing Problems
von: Levin, Joshua, et al.
Veröffentlicht: (2022)
von: Levin, Joshua, et al.
Veröffentlicht: (2022)
Federated Learning on Riemannian Manifolds: A Gradient-Free Projection-Based Approach
von: Wang, Hongye, et al.
Veröffentlicht: (2025)
von: Wang, Hongye, et al.
Veröffentlicht: (2025)
Policy Transfer for Continuous-Time Reinforcement Learning: A (Rough) Differential Equation Approach
von: Guo, Xin, et al.
Veröffentlicht: (2025)
von: Guo, Xin, et al.
Veröffentlicht: (2025)
Faster Gradient-Free Algorithms for Nonsmooth Nonconvex Stochastic Optimization
von: Chen, Lesi, et al.
Veröffentlicht: (2023)
von: Chen, Lesi, et al.
Veröffentlicht: (2023)
The ADMM-PINNs Algorithmic Framework for Nonsmooth PDE-Constrained Optimization: A Deep Learning Approach
von: Song, Yongcun, et al.
Veröffentlicht: (2023)
von: Song, Yongcun, et al.
Veröffentlicht: (2023)
Operator Models for Continuous-Time Offline Reinforcement Learning
von: Hoischen, Nicolas, et al.
Veröffentlicht: (2025)
von: Hoischen, Nicolas, et al.
Veröffentlicht: (2025)
A Novel Hybrid Heuristic-Reinforcement Learning Optimization Approach for a Class of Railcar Shunting Problems
von: Zhao, Ruonan, et al.
Veröffentlicht: (2026)
von: Zhao, Ruonan, et al.
Veröffentlicht: (2026)
Sign-SGD via Parameter-Free Optimization
von: Medyakov, Daniil, et al.
Veröffentlicht: (2025)
von: Medyakov, Daniil, et al.
Veröffentlicht: (2025)
Parameter-Free Algorithms for Performative Regret Minimization under Decision-Dependent Distributions
von: Park, Sungwoo, et al.
Veröffentlicht: (2024)
von: Park, Sungwoo, et al.
Veröffentlicht: (2024)
Completely Parameter-Free Single-Loop Algorithms for Nonconvex-Concave Minimax Problems
von: Yang, Junnan, et al.
Veröffentlicht: (2024)
von: Yang, Junnan, et al.
Veröffentlicht: (2024)
Parameter-Free Non-Ergodic Extragradient Algorithms for Solving Monotone Variational Inequalities
von: Shen, Lingqing, et al.
Veröffentlicht: (2026)
von: Shen, Lingqing, et al.
Veröffentlicht: (2026)
Precise Insulin Delivery for Artificial Pancreas: A Reinforcement Learning Optimized Adaptive Fuzzy Control Approach
von: Mameche, Omar, et al.
Veröffentlicht: (2025)
von: Mameche, Omar, et al.
Veröffentlicht: (2025)
From Learning to Optimize to Learning Optimization Algorithms
von: Castera, Camille, et al.
Veröffentlicht: (2024)
von: Castera, Camille, et al.
Veröffentlicht: (2024)
Accelerating Cutting-Plane Algorithms via Reinforcement Learning Surrogates
von: Mana, Kyle, et al.
Veröffentlicht: (2023)
von: Mana, Kyle, et al.
Veröffentlicht: (2023)
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
von: Ye, Lintao, et al.
Veröffentlicht: (2024)
von: Ye, Lintao, et al.
Veröffentlicht: (2024)
Gradient-Free Approaches is a Key to an Efficient Interaction with Markovian Stochasticity
von: Prokhorov, Boris, et al.
Veröffentlicht: (2026)
von: Prokhorov, Boris, et al.
Veröffentlicht: (2026)
Resilient Constrained Reinforcement Learning
von: Ding, Dongsheng, et al.
Veröffentlicht: (2023)
von: Ding, Dongsheng, et al.
Veröffentlicht: (2023)
Faster Adaptive Decentralized Learning Algorithms
von: Huang, Feihu, et al.
Veröffentlicht: (2024)
von: Huang, Feihu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
High-Order Error Bounds for Markovian LSA with Richardson-Romberg Extrapolation
von: Levin, Ilya, et al.
Veröffentlicht: (2025) -
Nonasymptotic Analysis of Stochastic Gradient Descent with the Richardson-Romberg Extrapolation
von: Sheshukova, Marina, et al.
Veröffentlicht: (2024) -
SCAFFLSA: Taming Heterogeneity in Federated Linear Stochastic Approximation and TD Learning
von: Mangold, Paul, et al.
Veröffentlicht: (2024) -
Statistical analysis of Inverse Entropy-regularized Reinforcement Learning
von: Belomestny, Denis, et al.
Veröffentlicht: (2025) -
Improved High-Probability Bounds for the Temporal Difference Learning Algorithm via Exponential Stability
von: Samsonov, Sergey, et al.
Veröffentlicht: (2023)