Off-policy estimation with adaptively collected data: the power of online learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Jeonghwan, Ma, Cong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
High-probability sample complexities for policy evaluation with linear function approximation
di: Li, Gen, et al.
Pubblicazione: (2023)
di: Li, Gen, et al.
Pubblicazione: (2023)
Learning bounds for doubly-robust covariate shift adaptation
di: Lee, Jeonghwan, et al.
Pubblicazione: (2025)
di: Lee, Jeonghwan, et al.
Pubblicazione: (2025)
$K$-Nearest-Neighbor Resampling for Off-Policy Evaluation in Stochastic Control
di: Giegrich, Michael, et al.
Pubblicazione: (2023)
di: Giegrich, Michael, et al.
Pubblicazione: (2023)
Risk reversal for least squares estimators under nested convex constraints
di: Al-Ghattas, Omar
Pubblicazione: (2026)
di: Al-Ghattas, Omar
Pubblicazione: (2026)
Iteratively reweighted kernel machines efficiently learn sparse functions
di: Zhu, Libin, et al.
Pubblicazione: (2025)
di: Zhu, Libin, et al.
Pubblicazione: (2025)
High-dimensional scaling limits and fluctuations of online least-squares SGD with smooth covariance
di: Balasubramanian, Krishnakumar, et al.
Pubblicazione: (2023)
di: Balasubramanian, Krishnakumar, et al.
Pubblicazione: (2023)
Joint learning of a network of linear dynamical systems via total variation penalization
di: Donnat, Claire, et al.
Pubblicazione: (2025)
di: Donnat, Claire, et al.
Pubblicazione: (2025)
Continuous-time reinforcement learning: ellipticity enables model-free value function approximation
di: Mou, Wenlong
Pubblicazione: (2026)
di: Mou, Wenlong
Pubblicazione: (2026)
Statistical guarantees for continuous-time policy evaluation: blessing of ellipticity and new tradeoffs
di: Mou, Wenlong
Pubblicazione: (2025)
di: Mou, Wenlong
Pubblicazione: (2025)
RandALO: Out-of-sample risk estimation in no time flat
di: Nobel, Parth, et al.
Pubblicazione: (2024)
di: Nobel, Parth, et al.
Pubblicazione: (2024)
Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models
di: Mou, Wenlong
Pubblicazione: (2025)
di: Mou, Wenlong
Pubblicazione: (2025)
Beyond Maximum Likelihood: Variational Inequality Estimation for Generalized Linear Models
di: Zhu, Linglingzhi, et al.
Pubblicazione: (2025)
di: Zhu, Linglingzhi, et al.
Pubblicazione: (2025)
Pessimism Meets Risk: Risk-Sensitive Offline Reinforcement Learning
di: Zhang, Dake, et al.
Pubblicazione: (2024)
di: Zhang, Dake, et al.
Pubblicazione: (2024)
Denoising Diffusions with Optimal Transport: Localization, Curvature, and Multi-Scale Complexity
di: Liang, Tengyuan, et al.
Pubblicazione: (2024)
di: Liang, Tengyuan, et al.
Pubblicazione: (2024)
Online Experimental Design With Estimation-Regret Trade-off Under Network Interference
di: Zhang, Zhiheng, et al.
Pubblicazione: (2024)
di: Zhang, Zhiheng, et al.
Pubblicazione: (2024)
On the Uniform Convergence of Subdifferentials in Stochastic Optimization and Learning
di: Ruan, Feng
Pubblicazione: (2024)
di: Ruan, Feng
Pubblicazione: (2024)
Implicit Regularization for Tubal Tensor Factorizations via Gradient Descent
di: Karnik, Santhosh, et al.
Pubblicazione: (2024)
di: Karnik, Santhosh, et al.
Pubblicazione: (2024)
The Collusion of Memory and Nonlinearity in Stochastic Approximation With Constant Stepsize
di: Huo, Dongyan, et al.
Pubblicazione: (2024)
di: Huo, Dongyan, et al.
Pubblicazione: (2024)
On Regularization via Early Stopping for Least Squares Regression
di: Sonthalia, Rishi, et al.
Pubblicazione: (2024)
di: Sonthalia, Rishi, et al.
Pubblicazione: (2024)
Probabilistic Guarantees of Stochastic Recursive Gradient in Non-Convex Finite Sum Problems
di: Zhong, Yanjie, et al.
Pubblicazione: (2024)
di: Zhong, Yanjie, et al.
Pubblicazione: (2024)
The High Line: Exact Risk and Learning Rate Curves of Stochastic Adaptive Learning Rate Algorithms
di: Collins-Woodfin, Elizabeth, et al.
Pubblicazione: (2024)
di: Collins-Woodfin, Elizabeth, et al.
Pubblicazione: (2024)
On the Sample Complexity of Set Membership Estimation for Linear Systems with Disturbances Bounded by Convex Sets
di: Xu, Haonan, et al.
Pubblicazione: (2024)
di: Xu, Haonan, et al.
Pubblicazione: (2024)
Hyperparameter tuning via trajectory predictions: Stochastic prox-linear methods in matrix sensing
di: Lou, Mengqi, et al.
Pubblicazione: (2024)
di: Lou, Mengqi, et al.
Pubblicazione: (2024)
An Improved Analysis of Langevin Algorithms with Prior Diffusion for Non-Log-Concave Sampling
di: Huang, Xunpeng, et al.
Pubblicazione: (2024)
di: Huang, Xunpeng, et al.
Pubblicazione: (2024)
Stochastic Optimization with Optimal Importance Sampling
di: Aolaritei, Liviu, et al.
Pubblicazione: (2025)
di: Aolaritei, Liviu, et al.
Pubblicazione: (2025)
A Theory of Feature Learning in Kernel Models
di: Chen, Yunlu, et al.
Pubblicazione: (2023)
di: Chen, Yunlu, et al.
Pubblicazione: (2023)
A review of NMF, PLSA, LBA, EMA, and LCA with a focus on the identifiability issue
di: Qi, Qianqian, et al.
Pubblicazione: (2025)
di: Qi, Qianqian, et al.
Pubblicazione: (2025)
High-dimensional Limit of SGD for Diagonal Linear Networks
di: Malaxechebarría, Begoña García, et al.
Pubblicazione: (2026)
di: Malaxechebarría, Begoña García, et al.
Pubblicazione: (2026)
Error Analysis of Triangular Optimal Transport Maps for Filtering
di: Al-Jarrah, Mohammad, et al.
Pubblicazione: (2025)
di: Al-Jarrah, Mohammad, et al.
Pubblicazione: (2025)
Online Inference of Constrained Optimization: Primal-Dual Optimality and Sequential Quadratic Programming
di: Gao, Yihang, et al.
Pubblicazione: (2025)
di: Gao, Yihang, et al.
Pubblicazione: (2025)
A Spectral Framework for Closed-Form Relative Density Estimation
di: Bach, Francis
Pubblicazione: (2026)
di: Bach, Francis
Pubblicazione: (2026)
Mixing Times and Privacy Analysis for the Projected Langevin Algorithm under a Modulus of Continuity
di: Bravo, Mario, et al.
Pubblicazione: (2025)
di: Bravo, Mario, et al.
Pubblicazione: (2025)
Learning and Decision-Making with Data: Optimal Formulations and Phase Transitions
di: Bennouna, Amine, et al.
Pubblicazione: (2021)
di: Bennouna, Amine, et al.
Pubblicazione: (2021)
Blessings and Curses of Covariate Shifts: Adversarial Learning Dynamics, Directional Convergence, and Equilibria
di: Liang, Tengyuan
Pubblicazione: (2022)
di: Liang, Tengyuan
Pubblicazione: (2022)
Extreme mass distributions for quasi-copulas
di: Omladič, Matjaž, et al.
Pubblicazione: (2025)
di: Omladič, Matjaž, et al.
Pubblicazione: (2025)
Learning an Optimal Assortment Policy under Observational Data
di: Han, Yuxuan, et al.
Pubblicazione: (2025)
di: Han, Yuxuan, et al.
Pubblicazione: (2025)
Convergence of flow-based generative models via proximal gradient descent in Wasserstein space
di: Cheng, Xiuyuan, et al.
Pubblicazione: (2023)
di: Cheng, Xiuyuan, et al.
Pubblicazione: (2023)
Algorithms for mean-field variational inference via polyhedral optimization in the Wasserstein space
di: Jiang, Yiheng, et al.
Pubblicazione: (2023)
di: Jiang, Yiheng, et al.
Pubblicazione: (2023)
An Elementary Proof of the Near Optimality of LogSumExp Smoothing
di: Samakhoana, Thabo, et al.
Pubblicazione: (2025)
di: Samakhoana, Thabo, et al.
Pubblicazione: (2025)
Federated Optimization of Smooth Loss Functions
di: Jadbabaie, Ali, et al.
Pubblicazione: (2022)
di: Jadbabaie, Ali, et al.
Pubblicazione: (2022)
Documenti analoghi
-
High-probability sample complexities for policy evaluation with linear function approximation
di: Li, Gen, et al.
Pubblicazione: (2023) -
Learning bounds for doubly-robust covariate shift adaptation
di: Lee, Jeonghwan, et al.
Pubblicazione: (2025) -
$K$-Nearest-Neighbor Resampling for Off-Policy Evaluation in Stochastic Control
di: Giegrich, Michael, et al.
Pubblicazione: (2023) -
Risk reversal for least squares estimators under nested convex constraints
di: Al-Ghattas, Omar
Pubblicazione: (2026) -
Iteratively reweighted kernel machines efficiently learn sparse functions
di: Zhu, Libin, et al.
Pubblicazione: (2025)