Multi-Agent Stage-wise Conservative Linear Bandits
Fuente:
arXiv
Saved in:
| Main Authors: | Afsharrad, Amirhossein, Moradipari, Ahmadreza, Lall, Sanjay |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cooperative Multi-Agent Constrained Stochastic Linear Bandits
by: Afsharrad, Amirhossein, et al.
Published: (2024)
by: Afsharrad, Amirhossein, et al.
Published: (2024)
On-Policy Distillation of Language Models for Autonomous Vehicle Motion Planning
by: Afsharrad, Amirhossein, et al.
Published: (2026)
by: Afsharrad, Amirhossein, et al.
Published: (2026)
One Goal, Many Challenges: Robust Preference Optimization Amid Content-Aware and Multi-Source Noise
by: Afzali, Amirabbas, et al.
Published: (2025)
by: Afzali, Amirabbas, et al.
Published: (2025)
Formation and Investigation of Cooperative Platooning at the Early Stage of Connected and Automated Vehicles Deployment
by: Mu, Zeyu, et al.
Published: (2025)
by: Mu, Zeyu, et al.
Published: (2025)
Learning to Sparsify Stochastic Linear Bandits
by: Wang, Zhengmiao, et al.
Published: (2026)
by: Wang, Zhengmiao, et al.
Published: (2026)
Finite-Time Guarantees for Multi-Agent Combinatorial Bandits with Nonstationary Rewards
by: Adams, Katherine B., et al.
Published: (2025)
by: Adams, Katherine B., et al.
Published: (2025)
Decentralized Upper Confidence Bound Algorithms for Homogeneous Multi-Agent Multi-Armed Bandits
by: Zhu, Jingxuan, et al.
Published: (2021)
by: Zhu, Jingxuan, et al.
Published: (2021)
Explore-then-Commit for Nonstationary Linear Bandits with Latent Dynamics
by: Choi, Sunmook, et al.
Published: (2025)
by: Choi, Sunmook, et al.
Published: (2025)
Beyond Binary Preferences: A Principled Framework for Reward Modeling with Ordinal Feedback
by: Afsharrad, Amirhossein, et al.
Published: (2026)
by: Afsharrad, Amirhossein, et al.
Published: (2026)
An Exploration-free Method for a Linear Stochastic Bandit Driven by a Linear Gaussian Dynamical System
by: Gornet, Jonathan, et al.
Published: (2025)
by: Gornet, Jonathan, et al.
Published: (2025)
Restless Bandit Problem with Rewards Generated by a Linear Gaussian Dynamical System
by: Gornet, Jonathan, et al.
Published: (2024)
by: Gornet, Jonathan, et al.
Published: (2024)
Tangential Randomization in Linear Bandits (TRAiL): Guaranteed Inference and Regret Bounds
by: Güçlü, Arda, et al.
Published: (2024)
by: Güçlü, Arda, et al.
Published: (2024)
Byzantine-Resilient Decentralized Multi-Armed Bandits
by: Zhu, Jingxuan, et al.
Published: (2023)
by: Zhu, Jingxuan, et al.
Published: (2023)
LORE: Lagrangian-Optimized Robust Embeddings for Visual Encoders
by: Khodabandeh, Borna, et al.
Published: (2025)
by: Khodabandeh, Borna, et al.
Published: (2025)
Multi-agent Multi-armed Bandits with Minimum Reward Guarantee Fairness
by: Manupriya, Piyushi, et al.
Published: (2025)
by: Manupriya, Piyushi, et al.
Published: (2025)
Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy
by: Juneja, Ishank, et al.
Published: (2026)
by: Juneja, Ishank, et al.
Published: (2026)
Semi-Supervised Multi-Task Learning Based Framework for Power System Security Assessment
by: Za'ter, Muhy Eddin, et al.
Published: (2024)
by: Za'ter, Muhy Eddin, et al.
Published: (2024)
Distributed Multi-Task Learning for Stochastic Bandits with Context Distribution and Stage-wise Constraints
by: Lin, Jiabin, et al.
Published: (2024)
by: Lin, Jiabin, et al.
Published: (2024)
Lagrangian Relaxation for Multi-Action Partially Observable Restless Bandits: Heuristic Policies and Indexability
by: Meshram, Rahul, et al.
Published: (2025)
by: Meshram, Rahul, et al.
Published: (2025)
Multi-User mmWave Beam and Rate Adaptation via Combinatorial Satisficing Bandits
by: Özyıldırım, Emre, et al.
Published: (2026)
by: Özyıldırım, Emre, et al.
Published: (2026)
An LP-based Sampling Policy for Multi-Armed Bandits with Side-Observations and Stochastic Availability
by: Soni, Ashutosh, et al.
Published: (2026)
by: Soni, Ashutosh, et al.
Published: (2026)
Bandit Algorithms for Deep Brain Stimulation
by: Gupta, Arkaprava, et al.
Published: (2026)
by: Gupta, Arkaprava, et al.
Published: (2026)
Causal Optimal Coupling for Gaussian Input-Output Distributional Data
by: Xu, Daran, et al.
Published: (2026)
by: Xu, Daran, et al.
Published: (2026)
Faster Q-Learning Algorithms for Restless Bandits
by: Kakarapalli, Parvish, et al.
Published: (2024)
by: Kakarapalli, Parvish, et al.
Published: (2024)
Adversarial Training of Two-Layer Polynomial and ReLU Activation Networks via Convex Optimization
by: Kuelbs, Daniel, et al.
Published: (2024)
by: Kuelbs, Daniel, et al.
Published: (2024)
Adversarial Network Optimization under Bandit Feedback: Maximizing Utility in Non-Stationary Multi-Hop Networks
by: Dai, Yan, et al.
Published: (2024)
by: Dai, Yan, et al.
Published: (2024)
Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes
by: Mittal, Vishesh, et al.
Published: (2024)
by: Mittal, Vishesh, et al.
Published: (2024)
Optimization and Learning in Open Multi-Agent Systems
by: Deplano, Diego, et al.
Published: (2025)
by: Deplano, Diego, et al.
Published: (2025)
Modeling Buffer Occupancy in bittide Systems
by: Lall, Sanjay, et al.
Published: (2024)
by: Lall, Sanjay, et al.
Published: (2024)
Reinforcement Learning-based Control via Y-wise Affine Neural Networks (YANNs)
by: Braniff, Austin, et al.
Published: (2025)
by: Braniff, Austin, et al.
Published: (2025)
Lyapunov Function Consistent Adaptive Network Signal Control with Back Pressure and Reinforcement Learning
by: Ma, Chaolun, et al.
Published: (2022)
by: Ma, Chaolun, et al.
Published: (2022)
FDM-Bench: A Comprehensive Benchmark for Evaluating Large Language Models in Additive Manufacturing Tasks
by: Eslaminia, Ahmadreza, et al.
Published: (2024)
by: Eslaminia, Ahmadreza, et al.
Published: (2024)
Risk-Aware Decision Making in Restless Bandits: Theory and Algorithms for Planning and Learning
by: Akbarzadeh, Nima, et al.
Published: (2024)
by: Akbarzadeh, Nima, et al.
Published: (2024)
Optimism as Risk-Seeking in Multi-Agent Reinforcement Learning
by: Zhang, Runyu, et al.
Published: (2025)
by: Zhang, Runyu, et al.
Published: (2025)
Regret Analysis of Multi-task Representation Learning for Linear-Quadratic Adaptive Control
by: Lee, Bruce D., et al.
Published: (2024)
by: Lee, Bruce D., et al.
Published: (2024)
Analysis of Off-Policy Multi-Step TD-Learning with Linear Function Approximation
by: Lee, Donghwan
Published: (2024)
by: Lee, Donghwan
Published: (2024)
Differentially Private High Dimensional Bandits
by: Shukla, Apurv
Published: (2024)
by: Shukla, Apurv
Published: (2024)
Logarithmically Quantized Distributed Optimization over Dynamic Multi-Agent Networks
by: Doostmohammadian, Mohammadreza, et al.
Published: (2024)
by: Doostmohammadian, Mohammadreza, et al.
Published: (2024)
Reinforcement Learning-based Control via Y-wise Affine Neural Networks: Comparative Case Studies for Chemical Processes
by: Braniff, Austin, et al.
Published: (2026)
by: Braniff, Austin, et al.
Published: (2026)
Structured Cooperative Multi-Agent Reinforcement Learning: a Bayesian Network Perspective
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
Similar Items
-
Cooperative Multi-Agent Constrained Stochastic Linear Bandits
by: Afsharrad, Amirhossein, et al.
Published: (2024) -
On-Policy Distillation of Language Models for Autonomous Vehicle Motion Planning
by: Afsharrad, Amirhossein, et al.
Published: (2026) -
One Goal, Many Challenges: Robust Preference Optimization Amid Content-Aware and Multi-Source Noise
by: Afzali, Amirabbas, et al.
Published: (2025) -
Formation and Investigation of Cooperative Platooning at the Early Stage of Connected and Automated Vehicles Deployment
by: Mu, Zeyu, et al.
Published: (2025) -
Learning to Sparsify Stochastic Linear Bandits
by: Wang, Zhengmiao, et al.
Published: (2026)