Nonstationary Reinforcement Learning with Linear Function Approximation
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhou, Huozhi, Chen, Jinglin, Varshney, Lav R., Jagmohan, Ashish |
|---|---|
| Formato: | Preprint |
| Publicado: |
2020
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Online Reinforcement Learning with Passive Memory
por: Pattanaik, Anay, et al.
Publicado: (2024)
por: Pattanaik, Anay, et al.
Publicado: (2024)
SparseJEPA: Sparse Representation Learning of Joint Embedding Predictive Architectures
por: Hartman, Max, et al.
Publicado: (2025)
por: Hartman, Max, et al.
Publicado: (2025)
A Meta-Learning Perspective on Transformers for Causal Language Modeling
por: Wu, Xinbo, et al.
Publicado: (2023)
por: Wu, Xinbo, et al.
Publicado: (2023)
Federated Learning via Lattice Joint Source-Channel Coding
por: Azimi-Abarghouyi, Seyed Mohammad, et al.
Publicado: (2024)
por: Azimi-Abarghouyi, Seyed Mohammad, et al.
Publicado: (2024)
Compute-Update Federated Learning: A Lattice Coding Approach Over-the-Air
por: Azimi-Abarghouyi, Seyed Mohammad, et al.
Publicado: (2024)
por: Azimi-Abarghouyi, Seyed Mohammad, et al.
Publicado: (2024)
Replicable Reinforcement Learning with Linear Function Approximation
por: Eaton, Eric, et al.
Publicado: (2025)
por: Eaton, Eric, et al.
Publicado: (2025)
Optimal Guarantees for Auditing Rényi Differentially Private Machine Learning
por: Kim, Benjamin D., et al.
Publicado: (2026)
por: Kim, Benjamin D., et al.
Publicado: (2026)
Reinforcement Learning with Function Approximation: From Linear to Nonlinear
por: Long, Jihao, et al.
Publicado: (2023)
por: Long, Jihao, et al.
Publicado: (2023)
Efficient Model-Agnostic Multi-Group Equivariant Networks
por: Baltaji, Razan, et al.
Publicado: (2023)
por: Baltaji, Razan, et al.
Publicado: (2023)
CAETC: Causal Autoencoding and Treatment Conditioning for Counterfactual Estimation over Time
por: Nguyen, Nghia D., et al.
Publicado: (2026)
por: Nguyen, Nghia D., et al.
Publicado: (2026)
Federated Nonlinear System Identification
por: Tupe, Omkar, et al.
Publicado: (2025)
por: Tupe, Omkar, et al.
Publicado: (2025)
Continual Generalized Category Discovery: Learning and Forgetting from a Bayesian Perspective
por: Dai, Hao, et al.
Publicado: (2025)
por: Dai, Hao, et al.
Publicado: (2025)
Distributionally Robust Off-Dynamics Reinforcement Learning: Provable Efficiency with Linear Function Approximation
por: Liu, Zhishuai, et al.
Publicado: (2024)
por: Liu, Zhishuai, et al.
Publicado: (2024)
Gap-Dependent Bounds for Nearly Minimax Optimal Reinforcement Learning with Linear Function Approximation
por: Zhang, Haochen, et al.
Publicado: (2026)
por: Zhang, Haochen, et al.
Publicado: (2026)
Adaptive Linear Embedding for Nonstationary High-Dimensional Optimization
por: Wen, Yuejiang, et al.
Publicado: (2025)
por: Wen, Yuejiang, et al.
Publicado: (2025)
Beyond Pooling: Matching for Robust Generalization under Data Heterogeneity
por: Roy, Ayush, et al.
Publicado: (2026)
por: Roy, Ayush, et al.
Publicado: (2026)
The Role of Inherent Bellman Error in Offline Reinforcement Learning with Linear Function Approximation
por: Golowich, Noah, et al.
Publicado: (2024)
por: Golowich, Noah, et al.
Publicado: (2024)
Linear Function Approximation as a Computationally Efficient Method to Solve Classical Reinforcement Learning Challenges
por: Srikanth, Hari
Publicado: (2024)
por: Srikanth, Hari
Publicado: (2024)
A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search
por: Ellis-Mohr, Austin R., et al.
Publicado: (2025)
por: Ellis-Mohr, Austin R., et al.
Publicado: (2025)
Provably Efficient Reinforcement Learning with Multinomial Logit Function Approximation
por: Li, Long-Fei, et al.
Publicado: (2024)
por: Li, Long-Fei, et al.
Publicado: (2024)
Nonstationary Generalized Linear Bandits with Discounted Online Mirror Descent
por: Lee, Joongkyu, et al.
Publicado: (2026)
por: Lee, Joongkyu, et al.
Publicado: (2026)
ViRN: Variational Inference and Distribution Trilateration for Long-Tailed Continual Representation Learning
por: Dai, Hao, et al.
Publicado: (2025)
por: Dai, Hao, et al.
Publicado: (2025)
Towards Differentially Private Reinforcement Learning with General Function Approximation
por: He, Yi, et al.
Publicado: (2026)
por: He, Yi, et al.
Publicado: (2026)
Reinforcement Learning from Partial Observation: Linear Function Approximation with Provable Sample Efficiency
por: Cai, Qi, et al.
Publicado: (2022)
por: Cai, Qi, et al.
Publicado: (2022)
Learning When to Restart: Nonstationary Newsvendor from Uncensored to Censored Demand
por: Chen, Xin, et al.
Publicado: (2025)
por: Chen, Xin, et al.
Publicado: (2025)
Provable Risk-Sensitive Distributional Reinforcement Learning with General Function Approximation
por: Chen, Yu, et al.
Publicado: (2024)
por: Chen, Yu, et al.
Publicado: (2024)
Strategically Robust Multi-Agent Reinforcement Learning with Linear Function Approximation
por: Gonzales, Jake, et al.
Publicado: (2026)
por: Gonzales, Jake, et al.
Publicado: (2026)
Causal Temporal Representation Learning with Nonstationary Sparse Transition
por: Song, Xiangchen, et al.
Publicado: (2024)
por: Song, Xiangchen, et al.
Publicado: (2024)
Sample and Computationally Efficient Continuous-Time Reinforcement Learning with General Function Approximation
por: Zhao, Runze, et al.
Publicado: (2025)
por: Zhao, Runze, et al.
Publicado: (2025)
Adaptive Resolving Methods for Reinforcement Learning with Function Approximations
por: Jiang, Jiashuo, et al.
Publicado: (2025)
por: Jiang, Jiashuo, et al.
Publicado: (2025)
Online Robust Reinforcement Learning with General Function Approximation
por: Ghosh, Debamita, et al.
Publicado: (2025)
por: Ghosh, Debamita, et al.
Publicado: (2025)
Nonstationary Sparse Spectral Permanental Process
por: Sun, Zicheng, et al.
Publicado: (2024)
por: Sun, Zicheng, et al.
Publicado: (2024)
Statistical Inference for Temporal Difference Learning with Linear Function Approximation
por: Wu, Weichen, et al.
Publicado: (2024)
por: Wu, Weichen, et al.
Publicado: (2024)
Accelerated Distributional Temporal Difference Learning with Linear Function Approximation
por: Jin, Kaicheng, et al.
Publicado: (2025)
por: Jin, Kaicheng, et al.
Publicado: (2025)
Convergence of Distributionally Robust Q-Learning with Linear Function Approximation
por: Mandal, Saptarshi, et al.
Publicado: (2025)
por: Mandal, Saptarshi, et al.
Publicado: (2025)
Explore-then-Commit for Nonstationary Linear Bandits with Latent Dynamics
por: Choi, Sunmook, et al.
Publicado: (2025)
por: Choi, Sunmook, et al.
Publicado: (2025)
Corruption-Robust Offline Reinforcement Learning with General Function Approximation
por: Ye, Chenlu, et al.
Publicado: (2023)
por: Ye, Chenlu, et al.
Publicado: (2023)
Model-Based Reinforcement Learning with Multinomial Logistic Function Approximation
por: Hwang, Taehyun, et al.
Publicado: (2022)
por: Hwang, Taehyun, et al.
Publicado: (2022)
Randomized Exploration for Reinforcement Learning with Multinomial Logistic Function Approximation
por: Cho, Wooseong, et al.
Publicado: (2024)
por: Cho, Wooseong, et al.
Publicado: (2024)
RL as Regressor: A Reinforcement Learning Approach for Function Approximation
por: Huang, Yongchao
Publicado: (2025)
por: Huang, Yongchao
Publicado: (2025)
Ejemplares similares
-
Online Reinforcement Learning with Passive Memory
por: Pattanaik, Anay, et al.
Publicado: (2024) -
SparseJEPA: Sparse Representation Learning of Joint Embedding Predictive Architectures
por: Hartman, Max, et al.
Publicado: (2025) -
A Meta-Learning Perspective on Transformers for Causal Language Modeling
por: Wu, Xinbo, et al.
Publicado: (2023) -
Federated Learning via Lattice Joint Source-Channel Coding
por: Azimi-Abarghouyi, Seyed Mohammad, et al.
Publicado: (2024) -
Compute-Update Federated Learning: A Lattice Coding Approach Over-the-Air
por: Azimi-Abarghouyi, Seyed Mohammad, et al.
Publicado: (2024)