Minimax-Optimal Reward-Agnostic Exploration in Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Gen, Yan, Yuling, Chen, Yuxin, Fan, Jianqing |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Model-Based Reinforcement Learning for Offline Zero-Sum Markov Games
by: Yan, Yuling, et al.
Published: (2022)
by: Yan, Yuling, et al.
Published: (2022)
Is Q-Learning Minimax Optimal? A Tight Sample Complexity Analysis
by: Li, Gen, et al.
Published: (2021)
by: Li, Gen, et al.
Published: (2021)
Settling the Sample Complexity of Model-Based Offline Reinforcement Learning
by: Li, Gen, et al.
Published: (2022)
by: Li, Gen, et al.
Published: (2022)
Inference for Heteroskedastic PCA with Missing Data
by: Yan, Yuling, et al.
Published: (2021)
by: Yan, Yuling, et al.
Published: (2021)
Breaking the Sample Size Barrier in Model-Based Reinforcement Learning with a Generative Model
by: Li, Gen, et al.
Published: (2020)
by: Li, Gen, et al.
Published: (2020)
Minimax Optimality of the Probability Flow ODE for Diffusion Models
by: Cai, Changxiao, et al.
Published: (2025)
by: Cai, Changxiao, et al.
Published: (2025)
The Curious Price of Distributional Robustness in Reinforcement Learning with a Generative Model
by: Shi, Laixi, et al.
Published: (2023)
by: Shi, Laixi, et al.
Published: (2023)
Towards Efficient Online Exploration for Reinforcement Learning with Human Feedback
by: Li, Gen, et al.
Published: (2025)
by: Li, Gen, et al.
Published: (2025)
Optimal transport natural gradient for statistical manifolds with continuous sample space
by: Chen, Yifan, et al.
Published: (2018)
by: Chen, Yifan, et al.
Published: (2018)
Towards Faster Non-Asymptotic Convergence for Diffusion-Based Generative Models
by: Li, Gen, et al.
Published: (2023)
by: Li, Gen, et al.
Published: (2023)
On Instability of Minimax Optimal Optimism-Based Bandit Algorithms
by: Praharaj, Samya, et al.
Published: (2025)
by: Praharaj, Samya, et al.
Published: (2025)
High-probability sample complexities for policy evaluation with linear function approximation
by: Li, Gen, et al.
Published: (2023)
by: Li, Gen, et al.
Published: (2023)
Minimax Optimality of Score-based Diffusion Models: Beyond the Density Lower Bound Assumptions
by: Zhang, Kaihong, et al.
Published: (2024)
by: Zhang, Kaihong, et al.
Published: (2024)
Span-Agnostic Optimal Sample Complexity and Oracle Inequalities for Average-Reward RL
by: Zurek, Matthew, et al.
Published: (2025)
by: Zurek, Matthew, et al.
Published: (2025)
Asymptotic Theory of Eigenvectors for Latent Embeddings with Generalized Laplacian Matrices
by: Fan, Jianqing, et al.
Published: (2025)
by: Fan, Jianqing, et al.
Published: (2025)
Optimal training-conditional regret for online conformal prediction
by: Liang, Jiadong, et al.
Published: (2026)
by: Liang, Jiadong, et al.
Published: (2026)
Rate-Optimal Non-Asymptotics for the Quadratic Prediction Error Method
by: Stamouli, Charis, et al.
Published: (2024)
by: Stamouli, Charis, et al.
Published: (2024)
When can weak latent factors be statistically inferred?
by: Fan, Jianqing, et al.
Published: (2024)
by: Fan, Jianqing, et al.
Published: (2024)
Differentially Private Distributed Estimation and Learning
by: Papachristou, Marios, et al.
Published: (2023)
by: Papachristou, Marios, et al.
Published: (2023)
Breaking AR's Sampling Bottleneck: Provable Acceleration via Diffusion Language Models
by: Li, Gen, et al.
Published: (2025)
by: Li, Gen, et al.
Published: (2025)
Spectral Ranking Inferences based on General Multiway Comparisons
by: Fan, Jianqing, et al.
Published: (2023)
by: Fan, Jianqing, et al.
Published: (2023)
Minimax Hypothesis Testing for the Bradley-Terry-Luce Model
by: Makur, Anuran, et al.
Published: (2024)
by: Makur, Anuran, et al.
Published: (2024)
Minimax optimal differentially private synthetic data for smooth queries
by: Ding, Rundong, et al.
Published: (2026)
by: Ding, Rundong, et al.
Published: (2026)
Minimax optimal submatrix detection: Sharp non-asymptotic rates
by: Knight, Parker, et al.
Published: (2026)
by: Knight, Parker, et al.
Published: (2026)
Isotonic Mechanism for Exponential Family Estimation in Machine Learning Peer Review
by: Yan, Yuling, et al.
Published: (2023)
by: Yan, Yuling, et al.
Published: (2023)
Towards a mathematical theory for consistency training in diffusion models
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Agnostic Sample Compression Schemes for Regression
by: Attias, Idan, et al.
Published: (2018)
by: Attias, Idan, et al.
Published: (2018)
O(d/T) Convergence Theory for Diffusion Probabilistic Models under Minimal Assumptions
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Adapting to Unknown Low-Dimensional Structures in Score-Based Diffusion Models
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Structure-Enhanced Deep Reinforcement Learning for Optimal Transmission Scheduling
by: Chen, Jiazheng, et al.
Published: (2022)
by: Chen, Jiazheng, et al.
Published: (2022)
Minimax Optimal Algorithms with Fixed-$k$-Nearest Neighbors
by: Ryu, J. Jon, et al.
Published: (2022)
by: Ryu, J. Jon, et al.
Published: (2022)
Statistical and Algorithmic Foundations of Reinforcement Learning
by: Chi, Yuejie, et al.
Published: (2025)
by: Chi, Yuejie, et al.
Published: (2025)
Learning linear dynamical systems under convex constraints
by: Tyagi, Hemant, et al.
Published: (2023)
by: Tyagi, Hemant, et al.
Published: (2023)
Joint Learning of Linear Dynamical Systems under Smoothness Constraints
by: Tyagi, Hemant
Published: (2024)
by: Tyagi, Hemant
Published: (2024)
Long-time dynamics and universality of nonconvex gradient descent
by: Han, Qiyang
Published: (2025)
by: Han, Qiyang
Published: (2025)
Mixing Time of the Proximal Sampler in Relative Fisher Information via Strong Data Processing Inequality
by: Wibisono, Andre
Published: (2025)
by: Wibisono, Andre
Published: (2025)
A Score-Based Density Formula, with Applications in Diffusion Generative Models
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Ensemble-Conditional Gaussian Processes (Ens-CGP): Representation, Geometry, and Inference
by: Ravela, Sai, et al.
Published: (2026)
by: Ravela, Sai, et al.
Published: (2026)
Kernel Mean Embedding Topology: Weak and Strong Forms for Stochastic Kernels and Implications for Model Learning
by: Saldi, Naci, et al.
Published: (2025)
by: Saldi, Naci, et al.
Published: (2025)
Markov Kernels, Distances and Optimal Control: A Parable of Linear Quadratic Non-Gaussian Distribution Steering
by: Teter, Alexis M. H., et al.
Published: (2025)
by: Teter, Alexis M. H., et al.
Published: (2025)
Similar Items
-
Model-Based Reinforcement Learning for Offline Zero-Sum Markov Games
by: Yan, Yuling, et al.
Published: (2022) -
Is Q-Learning Minimax Optimal? A Tight Sample Complexity Analysis
by: Li, Gen, et al.
Published: (2021) -
Settling the Sample Complexity of Model-Based Offline Reinforcement Learning
by: Li, Gen, et al.
Published: (2022) -
Inference for Heteroskedastic PCA with Missing Data
by: Yan, Yuling, et al.
Published: (2021) -
Breaking the Sample Size Barrier in Model-Based Reinforcement Learning with a Generative Model
by: Li, Gen, et al.
Published: (2020)