Distributed Thompson sampling under constrained communication
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zerefa, Saba, Ren, Zhaolin, Ma, Haitong, Li, Na |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Stochastic Nonlinear Control via Finite-dimensional Spectral Dynamic Embedding
von: Ren, Zhaolin, et al.
Veröffentlicht: (2023)
von: Ren, Zhaolin, et al.
Veröffentlicht: (2023)
Scalable spectral representations for multi-agent reinforcement learning in network MDPs
von: Ren, Zhaolin, et al.
Veröffentlicht: (2024)
von: Ren, Zhaolin, et al.
Veröffentlicht: (2024)
TS-RSR: A provably efficient approach for batch Bayesian Optimization
von: Ren, Zhaolin, et al.
Veröffentlicht: (2024)
von: Ren, Zhaolin, et al.
Veröffentlicht: (2024)
Alignment of large language models with constrained learning
von: Zhang, Botong, et al.
Veröffentlicht: (2025)
von: Zhang, Botong, et al.
Veröffentlicht: (2025)
Efficient Duple Perturbation Robustness in Low-rank MDPs
von: Hu, Yang, et al.
Veröffentlicht: (2024)
von: Hu, Yang, et al.
Veröffentlicht: (2024)
Communication-Efficient Stochastic Distributed Learning
von: Ren, Xiaoxing, et al.
Veröffentlicht: (2025)
von: Ren, Xiaoxing, et al.
Veröffentlicht: (2025)
Jointly Computation- and Communication-Efficient Distributed Learning
von: Ren, Xiaoxing, et al.
Veröffentlicht: (2025)
von: Ren, Xiaoxing, et al.
Veröffentlicht: (2025)
Optimism as Risk-Seeking in Multi-Agent Reinforcement Learning
von: Zhang, Runyu, et al.
Veröffentlicht: (2025)
von: Zhang, Runyu, et al.
Veröffentlicht: (2025)
Distributed Online Submodular Maximization under Communication Delays: A Simultaneous Decision-Making Approach
von: Xu, Zirui, et al.
Veröffentlicht: (2026)
von: Xu, Zirui, et al.
Veröffentlicht: (2026)
Analysis of Thompson Sampling for Controlling Unknown Linear Diffusion Processes
von: Faradonbeh, Mohamad Kazem Shirani, et al.
Veröffentlicht: (2022)
von: Faradonbeh, Mohamad Kazem Shirani, et al.
Veröffentlicht: (2022)
Bayesian dynamic scheduling of multipurpose batch processes under incomplete look-ahead information
von: Zheng, Taicheng, et al.
Veröffentlicht: (2025)
von: Zheng, Taicheng, et al.
Veröffentlicht: (2025)
Convergence and sample complexity of natural policy gradient primal-dual methods for constrained MDPs
von: Ding, Dongsheng, et al.
Veröffentlicht: (2022)
von: Ding, Dongsheng, et al.
Veröffentlicht: (2022)
On the Foundation of Distributionally Robust Reinforcement Learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
MPC-Inspired Reinforcement Learning for Verifiable Model-Free Control
von: Lu, Yiwen, et al.
Veröffentlicht: (2023)
von: Lu, Yiwen, et al.
Veröffentlicht: (2023)
Modular Distributed Nonconvex Learning with Error Feedback
von: Carnevale, Guido, et al.
Veröffentlicht: (2025)
von: Carnevale, Guido, et al.
Veröffentlicht: (2025)
Robust Q-Learning under Corrupted Rewards
von: Maity, Sreejeet, et al.
Veröffentlicht: (2024)
von: Maity, Sreejeet, et al.
Veröffentlicht: (2024)
Decision-Dependent Stochastic Optimization: The Role of Distribution Dynamics
von: He, Zhiyu, et al.
Veröffentlicht: (2025)
von: He, Zhiyu, et al.
Veröffentlicht: (2025)
Online Control of Linear Systems under Unbounded Noise
von: Ito, Kaito, et al.
Veröffentlicht: (2024)
von: Ito, Kaito, et al.
Veröffentlicht: (2024)
Enhancing Convergence of Decentralized Gradient Tracking under the KL Property
von: Chen, Xiaokai, et al.
Veröffentlicht: (2024)
von: Chen, Xiaokai, et al.
Veröffentlicht: (2024)
Robust Neural IDA-PBC: passivity-based stabilization under approximations
von: Sanchez-Escalonilla, Santiago, et al.
Veröffentlicht: (2024)
von: Sanchez-Escalonilla, Santiago, et al.
Veröffentlicht: (2024)
On Linear Convergence of PI Consensus Algorithm under the Restricted Secant Inequality
von: Chakrabarti, Kushal, et al.
Veröffentlicht: (2023)
von: Chakrabarti, Kushal, et al.
Veröffentlicht: (2023)
Adversarially and Distributionally Robust Virtual Energy Storage Systems via the Scenario Approach
von: Pantazis, Georgios, et al.
Veröffentlicht: (2025)
von: Pantazis, Georgios, et al.
Veröffentlicht: (2025)
Data-driven Reachable Set Estimation with Tunable Adversarial and Wasserstein Distributional Guarantees
von: Pantazis, Georgios, et al.
Veröffentlicht: (2026)
von: Pantazis, Georgios, et al.
Veröffentlicht: (2026)
Distributed Control of Network Systems in the Space of Stabilizing Graph Neural Network Policies
von: Cao, John, et al.
Veröffentlicht: (2025)
von: Cao, John, et al.
Veröffentlicht: (2025)
Distributionally Robust Regret Optimal Control Under Moment-Based Ambiguity Sets
von: Taha, Feras Al, et al.
Veröffentlicht: (2025)
von: Taha, Feras Al, et al.
Veröffentlicht: (2025)
High-Probability Bounds for SGD under the Polyak-Lojasiewicz Condition with Markovian Noise
von: Kar, Avik, et al.
Veröffentlicht: (2026)
von: Kar, Avik, et al.
Veröffentlicht: (2026)
Sampling-Horizon Neural Operator Predictors for Nonlinear Control under Delayed Inputs
von: Bhan, Luke, et al.
Veröffentlicht: (2026)
von: Bhan, Luke, et al.
Veröffentlicht: (2026)
Wasserstein Distributionally Robust Nash Equilibrium Seeking with Heterogeneous Data: A Lagrangian Approach
von: Wang, Zifan, et al.
Veröffentlicht: (2025)
von: Wang, Zifan, et al.
Veröffentlicht: (2025)
Differentiable Distributionally Robust Optimization Layers
von: Ma, Xutao, et al.
Veröffentlicht: (2024)
von: Ma, Xutao, et al.
Veröffentlicht: (2024)
PolyFormer: learning efficient reformulations for scalable optimization under complex physical constraints
von: Wen, Yilin, et al.
Veröffentlicht: (2026)
von: Wen, Yilin, et al.
Veröffentlicht: (2026)
Robustness of Iteratively Pre-Conditioned Gradient-Descent Method: The Case of Distributed Linear Regression Problem
von: Chakrabarti, Kushal, et al.
Veröffentlicht: (2021)
von: Chakrabarti, Kushal, et al.
Veröffentlicht: (2021)
Iterative Pre-Conditioning for Expediting the Gradient-Descent Method: The Distributed Linear Least-Squares Problem
von: Chakrabarti, Kushal, et al.
Veröffentlicht: (2020)
von: Chakrabarti, Kushal, et al.
Veröffentlicht: (2020)
Evaluation of Prosumer Networks for Peak Load Management in Iran: A Distributed Contextual Stochastic Optimization Approach
von: Noori, Amir, et al.
Veröffentlicht: (2024)
von: Noori, Amir, et al.
Veröffentlicht: (2024)
Distributed online constrained convex optimization with event-triggered communication
von: Zhang, Kunpeng, et al.
Veröffentlicht: (2023)
von: Zhang, Kunpeng, et al.
Veröffentlicht: (2023)
An L-BFGS-B approach for linear and nonlinear system identification under $\ell_1$ and group-Lasso regularization
von: Bemporad, Alberto
Veröffentlicht: (2024)
von: Bemporad, Alberto
Veröffentlicht: (2024)
Convex Chance-Constrained Stochastic Control under Uncertain Specifications with Application to Learning-Based Hybrid Powertrain Control
von: Kato, Teruki, et al.
Veröffentlicht: (2026)
von: Kato, Teruki, et al.
Veröffentlicht: (2026)
Near-Optimal Distributed Linear-Quadratic Regulator for Networked Systems
von: Shin, Sungho, et al.
Veröffentlicht: (2022)
von: Shin, Sungho, et al.
Veröffentlicht: (2022)
Slack More, Predict Better: Proximal Relaxation for Probabilistic Latent Variable Model-based Soft Sensors
von: Zou, Zehua, et al.
Veröffentlicht: (2026)
von: Zou, Zehua, et al.
Veröffentlicht: (2026)
Logarithmically Quantized Distributed Optimization over Dynamic Multi-Agent Networks
von: Doostmohammadian, Mohammadreza, et al.
Veröffentlicht: (2024)
von: Doostmohammadian, Mohammadreza, et al.
Veröffentlicht: (2024)
Distributionally Robust Policy and Lyapunov-Certificate Learning
von: Long, Kehan, et al.
Veröffentlicht: (2024)
von: Long, Kehan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Stochastic Nonlinear Control via Finite-dimensional Spectral Dynamic Embedding
von: Ren, Zhaolin, et al.
Veröffentlicht: (2023) -
Scalable spectral representations for multi-agent reinforcement learning in network MDPs
von: Ren, Zhaolin, et al.
Veröffentlicht: (2024) -
TS-RSR: A provably efficient approach for batch Bayesian Optimization
von: Ren, Zhaolin, et al.
Veröffentlicht: (2024) -
Alignment of large language models with constrained learning
von: Zhang, Botong, et al.
Veröffentlicht: (2025) -
Efficient Duple Perturbation Robustness in Low-rank MDPs
von: Hu, Yang, et al.
Veröffentlicht: (2024)