Predictive Control and Regret Analysis of Non-Stationary MDP with Look-ahead Information
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Ziyi, Nakahira, Yorie, Qu, Guannan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Stabilize Unknown LTI Systems on a Single Trajectory under Stochastic Noise
by: Zhang, Ziyi, et al.
Published: (2024)
by: Zhang, Ziyi, et al.
Published: (2024)
Polynomial Convergence of Riemannian Diffusion Models
by: Xu, Xingyu, et al.
Published: (2026)
by: Xu, Xingyu, et al.
Published: (2026)
Stabilizing Linear Systems under Partial Observability: Sample Complexity and Fundamental Limits
by: Zhang, Ziyi, et al.
Published: (2025)
by: Zhang, Ziyi, et al.
Published: (2025)
UAVGENT: A Language-Guided Distributed Control Framework
by: Zhang, Ziyi, et al.
Published: (2026)
by: Zhang, Ziyi, et al.
Published: (2026)
Physics-informed RL for Maximal Safety Probability Estimation
by: Hoshino, Hikaru, et al.
Published: (2024)
by: Hoshino, Hikaru, et al.
Published: (2024)
A Generalizable Physics-informed Learning Framework for Risk Probability Estimation
by: Wang, Zhuoyuan, et al.
Published: (2023)
by: Wang, Zhuoyuan, et al.
Published: (2023)
Generalizable Physics-Informed Learning for Stochastic Safety-Critical Systems
by: Wang, Zhuoyuan, et al.
Published: (2024)
by: Wang, Zhuoyuan, et al.
Published: (2024)
Myopically Verifiable Probabilistic Certificates for Safe Control and Learning
by: Wang, Zhuoyuan, et al.
Published: (2024)
by: Wang, Zhuoyuan, et al.
Published: (2024)
Locally Interdependent Multi-Agent MDP: Theoretical Framework for Decentralized Agents with Dynamic Dependencies
by: DeWeese, Alex, et al.
Published: (2024)
by: DeWeese, Alex, et al.
Published: (2024)
Natural Policy Gradient for Average Reward Non-Stationary RL
by: Jali, Neharika, et al.
Published: (2025)
by: Jali, Neharika, et al.
Published: (2025)
An Analytic Solution to Covariance Propagation in Neural Networks
by: Wright, Oren, et al.
Published: (2024)
by: Wright, Oren, et al.
Published: (2024)
Neural Spline Operators for Risk Quantification in Stochastic Systems
by: Wang, Zhuoyuan, et al.
Published: (2025)
by: Wang, Zhuoyuan, et al.
Published: (2025)
Physics-Informed Deep B-Spline Networks
by: Wang, Zhuoyuan, et al.
Published: (2025)
by: Wang, Zhuoyuan, et al.
Published: (2025)
Kalman Bayesian Transformer
by: Jing, Haoming, et al.
Published: (2025)
by: Jing, Haoming, et al.
Published: (2025)
OpInf-LLM: Parametric PDE Solving with LLMs via Operator Inference
by: Wang, Zhuoyuan, et al.
Published: (2026)
by: Wang, Zhuoyuan, et al.
Published: (2026)
Logarithmic Regret and Polynomial Scaling in Online Multi-step-ahead Prediction
by: Qian, Jiachen, et al.
Published: (2025)
by: Qian, Jiachen, et al.
Published: (2025)
Optimal Dynamic Regret by Transformers for Non-Stationary Reinforcement Learning
by: Chen, Baiyuan, et al.
Published: (2025)
by: Chen, Baiyuan, et al.
Published: (2025)
Improved Regret Analysis in Gaussian Process Bandits: Optimality for Noiseless Reward, RKHS norm, and Non-Stationary Variance
by: Iwazaki, Shogo, et al.
Published: (2025)
by: Iwazaki, Shogo, et al.
Published: (2025)
Diffuse-CLoC: Guided Diffusion for Physics-based Character Look-ahead Control
by: Huang, Xiaoyu, et al.
Published: (2025)
by: Huang, Xiaoyu, et al.
Published: (2025)
Catoni-Style Change Point Detection for Regret Minimization in Non-Stationary Heavy-Tailed Bandits
by: Genalti, Gianmarco, et al.
Published: (2025)
by: Genalti, Gianmarco, et al.
Published: (2025)
Geometric Re-Analysis of Classical MDP Solving Algorithms
by: Mustafin, Arsenii, et al.
Published: (2025)
by: Mustafin, Arsenii, et al.
Published: (2025)
A Fast and Effective Solution to the Problem of Look-ahead Bias in LLMs
by: Merchant, Humzah, et al.
Published: (2025)
by: Merchant, Humzah, et al.
Published: (2025)
Non-Stationary Bandit Learning via Predictive Sampling
by: Liu, Yueyang, et al.
Published: (2022)
by: Liu, Yueyang, et al.
Published: (2022)
Non-Stationary Online Structured Prediction with Surrogate Losses
by: Sakaue, Shinsaku, et al.
Published: (2025)
by: Sakaue, Shinsaku, et al.
Published: (2025)
Wavelet Predictive Representations for Non-Stationary Reinforcement Learning
by: Wang, Min, et al.
Published: (2025)
by: Wang, Min, et al.
Published: (2025)
Safety Certificate against Latent Variables with Partially Unidentifiable Dynamics
by: Jing, Haoming, et al.
Published: (2025)
by: Jing, Haoming, et al.
Published: (2025)
Revisiting Policy Gradients for Restricted Policy Classes: Escaping Myopic Local Optima with $k$-step Policy Gradients
by: DeWeese, Alex, et al.
Published: (2026)
by: DeWeese, Alex, et al.
Published: (2026)
MDP Planning as Policy Inference
by: Tolpin, David
Published: (2026)
by: Tolpin, David
Published: (2026)
Towards Effective Theory of LLMs: A Representation Learning Approach
by: Ustaomeroglu, Muhammed, et al.
Published: (2026)
by: Ustaomeroglu, Muhammed, et al.
Published: (2026)
Sample-Optimal Zero-Violation Safety For Continuous Control
by: Ray, Ritabrata, et al.
Published: (2024)
by: Ray, Ritabrata, et al.
Published: (2024)
Non-Stationary Inventory Control with Lead Times
by: Amiri, Nele H., et al.
Published: (2026)
by: Amiri, Nele H., et al.
Published: (2026)
Track-MDP: Reinforcement Learning for Target Tracking with Controlled Sensing
by: Subramaniam, Adarsh M., et al.
Published: (2024)
by: Subramaniam, Adarsh M., et al.
Published: (2024)
BALLAST: Bayesian Active Learning with Look-ahead Amendment for Sea-drifter Trajectories under Spatio-Temporal Vector Fields
by: Zhang, Rui-Yang, et al.
Published: (2025)
by: Zhang, Rui-Yang, et al.
Published: (2025)
Back to the Future: Look-ahead Augmentation and Parallel Self-Refinement for Time Series Forecasting
by: Kim, Sunho, et al.
Published: (2026)
by: Kim, Sunho, et al.
Published: (2026)
Hedging Memory Horizons for Non-Stationary Prediction via Online Aggregation
by: Wang, Yutong, et al.
Published: (2026)
by: Wang, Yutong, et al.
Published: (2026)
Non-Stationary Lipschitz Bandits
by: Nguyen, Nicolas, et al.
Published: (2025)
by: Nguyen, Nicolas, et al.
Published: (2025)
Efficient Reinforcement Learning for Global Decision Making in the Presence of Local Agents at Scale
by: Anand, Emile, et al.
Published: (2024)
by: Anand, Emile, et al.
Published: (2024)
Logarithmic Regret for Nonlinear Control
by: Wang, James, et al.
Published: (2025)
by: Wang, James, et al.
Published: (2025)
Learning to Spend: Model Predictive Control for Budgeting under Non-Stationary Returns
by: Pathak, Nilavra, et al.
Published: (2026)
by: Pathak, Nilavra, et al.
Published: (2026)
Federated Learning With Energy Harvesting Devices: An MDP Framework
by: Zhang, Kai, et al.
Published: (2024)
by: Zhang, Kai, et al.
Published: (2024)
Similar Items
-
Learning to Stabilize Unknown LTI Systems on a Single Trajectory under Stochastic Noise
by: Zhang, Ziyi, et al.
Published: (2024) -
Polynomial Convergence of Riemannian Diffusion Models
by: Xu, Xingyu, et al.
Published: (2026) -
Stabilizing Linear Systems under Partial Observability: Sample Complexity and Fundamental Limits
by: Zhang, Ziyi, et al.
Published: (2025) -
UAVGENT: A Language-Guided Distributed Control Framework
by: Zhang, Ziyi, et al.
Published: (2026) -
Physics-informed RL for Maximal Safety Probability Estimation
by: Hoshino, Hikaru, et al.
Published: (2024)