Byzantine-Resilient Decentralized Multi-Armed Bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Jingxuan, Koppel, Alec, Velasquez, Alvaro, Liu, Ji |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Decentralized Upper Confidence Bound Algorithms for Homogeneous Multi-Agent Multi-Armed Bandits
von: Zhu, Jingxuan, et al.
Veröffentlicht: (2021)
von: Zhu, Jingxuan, et al.
Veröffentlicht: (2021)
Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy
von: Juneja, Ishank, et al.
Veröffentlicht: (2026)
von: Juneja, Ishank, et al.
Veröffentlicht: (2026)
An LP-based Sampling Policy for Multi-Armed Bandits with Side-Observations and Stochastic Availability
von: Soni, Ashutosh, et al.
Veröffentlicht: (2026)
von: Soni, Ashutosh, et al.
Veröffentlicht: (2026)
Nonparametric Sparse Online Learning of the Koopman Operator
von: Hou, Boya, et al.
Veröffentlicht: (2025)
von: Hou, Boya, et al.
Veröffentlicht: (2025)
Convergence of Byzantine-Resilient Gradient Tracking via Probabilistic Edge Dropout
von: Dezhboro, Amirhossein, et al.
Veröffentlicht: (2026)
von: Dezhboro, Amirhossein, et al.
Veröffentlicht: (2026)
Byzantine-resilient federated online learning for Gaussian process regression
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
Robust Decentralized Multi-armed Bandits: From Corruption-Resilience to Byzantine-Resilience
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
Multi-Armed Bandits with Minimum Aggregated Revenue Constraints
von: Yahmed, Ahmed Ben, et al.
Veröffentlicht: (2025)
von: Yahmed, Ahmed Ben, et al.
Veröffentlicht: (2025)
Multi-agent Multi-armed Bandits with Minimum Reward Guarantee Fairness
von: Manupriya, Piyushi, et al.
Veröffentlicht: (2025)
von: Manupriya, Piyushi, et al.
Veröffentlicht: (2025)
Multi-Agent Stage-wise Conservative Linear Bandits
von: Afsharrad, Amirhossein, et al.
Veröffentlicht: (2025)
von: Afsharrad, Amirhossein, et al.
Veröffentlicht: (2025)
Decentralized Optimization on Compact Submanifolds by Quantized Riemannian Gradient Tracking
von: Chen, Jun, et al.
Veröffentlicht: (2025)
von: Chen, Jun, et al.
Veröffentlicht: (2025)
Federated Multi-Armed Bandits Under Byzantine Attacks
von: Saday, Artun, et al.
Veröffentlicht: (2022)
von: Saday, Artun, et al.
Veröffentlicht: (2022)
Learning to Sparsify Stochastic Linear Bandits
von: Wang, Zhengmiao, et al.
Veröffentlicht: (2026)
von: Wang, Zhengmiao, et al.
Veröffentlicht: (2026)
Enabling Clean Energy Resilience with Machine Learning-Empowered Underground Hydrogen Storage
von: Carbonero, Alvaro, et al.
Veröffentlicht: (2024)
von: Carbonero, Alvaro, et al.
Veröffentlicht: (2024)
Finite-Time Guarantees for Multi-Agent Combinatorial Bandits with Nonstationary Rewards
von: Adams, Katherine B., et al.
Veröffentlicht: (2025)
von: Adams, Katherine B., et al.
Veröffentlicht: (2025)
Multi-User mmWave Beam and Rate Adaptation via Combinatorial Satisficing Bandits
von: Özyıldırım, Emre, et al.
Veröffentlicht: (2026)
von: Özyıldırım, Emre, et al.
Veröffentlicht: (2026)
Lagrangian Relaxation for Multi-Action Partially Observable Restless Bandits: Heuristic Policies and Indexability
von: Meshram, Rahul, et al.
Veröffentlicht: (2025)
von: Meshram, Rahul, et al.
Veröffentlicht: (2025)
Adaptive Testing Environment Generation for Connected and Automated Vehicles with Dense Reinforcement Learning
von: Yang, Jingxuan, et al.
Veröffentlicht: (2024)
von: Yang, Jingxuan, et al.
Veröffentlicht: (2024)
Decentralized Event-Triggered Online Learning for Safe Consensus of Multi-Agent Systems with Gaussian Process Regression
von: Dai, Xiaobing, et al.
Veröffentlicht: (2024)
von: Dai, Xiaobing, et al.
Veröffentlicht: (2024)
Bandit Algorithms for Deep Brain Stimulation
von: Gupta, Arkaprava, et al.
Veröffentlicht: (2026)
von: Gupta, Arkaprava, et al.
Veröffentlicht: (2026)
Decentralized Riemannian Conjugate Gradient Method on the Stiefel Manifold
von: Chen, Jun, et al.
Veröffentlicht: (2023)
von: Chen, Jun, et al.
Veröffentlicht: (2023)
Faster Q-Learning Algorithms for Restless Bandits
von: Kakarapalli, Parvish, et al.
Veröffentlicht: (2024)
von: Kakarapalli, Parvish, et al.
Veröffentlicht: (2024)
Sharpened Lazy Incremental Quasi-Newton Method
von: Lahoti, Aakash, et al.
Veröffentlicht: (2023)
von: Lahoti, Aakash, et al.
Veröffentlicht: (2023)
A Hessian-Free Actor-Critic Algorithm for Bi-Level Reinforcement Learning with Applications to LLM Fine-Tuning
von: Zeng, Sihan, et al.
Veröffentlicht: (2026)
von: Zeng, Sihan, et al.
Veröffentlicht: (2026)
Explore-then-Commit for Nonstationary Linear Bandits with Latent Dynamics
von: Choi, Sunmook, et al.
Veröffentlicht: (2025)
von: Choi, Sunmook, et al.
Veröffentlicht: (2025)
A Variance-Reduced Stochastic Gradient Tracking Algorithm for Decentralized Optimization with Orthogonality Constraints
von: Wang, Lei, et al.
Veröffentlicht: (2022)
von: Wang, Lei, et al.
Veröffentlicht: (2022)
Adversarial Network Optimization under Bandit Feedback: Maximizing Utility in Non-Stationary Multi-Hop Networks
von: Dai, Yan, et al.
Veröffentlicht: (2024)
von: Dai, Yan, et al.
Veröffentlicht: (2024)
Hierarchical Decentralized Stochastic Control for Cyber-Physical Systems
von: Kaza, Kesav, et al.
Veröffentlicht: (2025)
von: Kaza, Kesav, et al.
Veröffentlicht: (2025)
Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes
von: Mittal, Vishesh, et al.
Veröffentlicht: (2024)
von: Mittal, Vishesh, et al.
Veröffentlicht: (2024)
Accelerating Optimization and Machine Learning through Decentralization
von: Chen, Ziqin, et al.
Veröffentlicht: (2026)
von: Chen, Ziqin, et al.
Veröffentlicht: (2026)
Learning Decentralized Linear Quadratic Regulators with $\sqrt{T}$ Regret
von: Ye, Lintao, et al.
Veröffentlicht: (2022)
von: Ye, Lintao, et al.
Veröffentlicht: (2022)
Resilient Constrained Reinforcement Learning
von: Ding, Dongsheng, et al.
Veröffentlicht: (2023)
von: Ding, Dongsheng, et al.
Veröffentlicht: (2023)
Priority-Driven Control and Communication in Decentralized Multi-Agent Systems via Reinforcement Learning
von: Guo, Qingyun, et al.
Veröffentlicht: (2026)
von: Guo, Qingyun, et al.
Veröffentlicht: (2026)
Enhancing Convergence of Decentralized Gradient Tracking under the KL Property
von: Chen, Xiaokai, et al.
Veröffentlicht: (2024)
von: Chen, Xiaokai, et al.
Veröffentlicht: (2024)
Restless Bandit Problem with Rewards Generated by a Linear Gaussian Dynamical System
von: Gornet, Jonathan, et al.
Veröffentlicht: (2024)
von: Gornet, Jonathan, et al.
Veröffentlicht: (2024)
Tangential Randomization in Linear Bandits (TRAiL): Guaranteed Inference and Regret Bounds
von: Güçlü, Arda, et al.
Veröffentlicht: (2024)
von: Güçlü, Arda, et al.
Veröffentlicht: (2024)
Risk-Aware Decision Making in Restless Bandits: Theory and Algorithms for Planning and Learning
von: Akbarzadeh, Nima, et al.
Veröffentlicht: (2024)
von: Akbarzadeh, Nima, et al.
Veröffentlicht: (2024)
Differentially Private High Dimensional Bandits
von: Shukla, Apurv
Veröffentlicht: (2024)
von: Shukla, Apurv
Veröffentlicht: (2024)
TAG: A Decentralized Framework for Multi-Agent Hierarchical Reinforcement Learning
von: Paolo, Giuseppe, et al.
Veröffentlicht: (2025)
von: Paolo, Giuseppe, et al.
Veröffentlicht: (2025)
Federated Combinatorial Multi-Agent Multi-Armed Bandits
von: Fourati, Fares, et al.
Veröffentlicht: (2024)
von: Fourati, Fares, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Decentralized Upper Confidence Bound Algorithms for Homogeneous Multi-Agent Multi-Armed Bandits
von: Zhu, Jingxuan, et al.
Veröffentlicht: (2021) -
Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy
von: Juneja, Ishank, et al.
Veröffentlicht: (2026) -
An LP-based Sampling Policy for Multi-Armed Bandits with Side-Observations and Stochastic Availability
von: Soni, Ashutosh, et al.
Veröffentlicht: (2026) -
Nonparametric Sparse Online Learning of the Koopman Operator
von: Hou, Boya, et al.
Veröffentlicht: (2025) -
Convergence of Byzantine-Resilient Gradient Tracking via Probabilistic Edge Dropout
von: Dezhboro, Amirhossein, et al.
Veröffentlicht: (2026)