Multi-Agent Lipschitz Bandits
Fuente:
arXiv
Saved in:
| Main Authors: | Chakraborty, Sourav, Rege, Amit Kiran, Monteleoni, Claire, Chen, Lijun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Incentivized Lipschitz Bandits
by: Chakraborty, Sourav, et al.
Published: (2025)
by: Chakraborty, Sourav, et al.
Published: (2025)
Flickering Multi-Armed Bandits
by: Chakraborty, Sourav, et al.
Published: (2026)
by: Chakraborty, Sourav, et al.
Published: (2026)
A Unified Framework for Locality in Scalable MARL
by: Chakraborty, Sourav, et al.
Published: (2026)
by: Chakraborty, Sourav, et al.
Published: (2026)
The Role of Generator Access in Autoregressive Post-Training
by: Rege, Amit Kiran
Published: (2026)
by: Rege, Amit Kiran
Published: (2026)
Data Attribution in Adaptive Learning
by: Rege, Amit Kiran
Published: (2026)
by: Rege, Amit Kiran
Published: (2026)
Where Did Your Model Learn That? Label-free Influence for Self-supervised Learning
by: Harilal, Nidhin, et al.
Published: (2024)
by: Harilal, Nidhin, et al.
Published: (2024)
Incentivized Exploration of Non-Stationary Stochastic Bandits
by: Chakraborty, Sourav, et al.
Published: (2024)
by: Chakraborty, Sourav, et al.
Published: (2024)
Non-Stationary Lipschitz Bandits
by: Nguyen, Nicolas, et al.
Published: (2025)
by: Nguyen, Nicolas, et al.
Published: (2025)
Deep Clustering via Probabilistic Ratio-Cut Optimization
by: Ghriss, Ayoub, et al.
Published: (2025)
by: Ghriss, Ayoub, et al.
Published: (2025)
Hamiltonian Learning using Machine Learning Models Trained with Continuous Measurements
by: Tucker, Kris, et al.
Published: (2024)
by: Tucker, Kris, et al.
Published: (2024)
Quantum Lipschitz Bandits
by: Yi, Bongsoo, et al.
Published: (2025)
by: Yi, Bongsoo, et al.
Published: (2025)
Generating ensembles of spatially-coherent in-situ forecasts using flow matching
by: Landry, David, et al.
Published: (2025)
by: Landry, David, et al.
Published: (2025)
Lipschitz Bandits with Stochastic Delayed Feedback
by: Liu, Zhongxuan, et al.
Published: (2025)
by: Liu, Zhongxuan, et al.
Published: (2025)
ArchesWeather: An efficient AI weather forecasting model at 1.5° resolution
by: Couairon, Guillaume, et al.
Published: (2024)
by: Couairon, Guillaume, et al.
Published: (2024)
On the Sample Complexity of Two-Layer Networks: Lipschitz vs. Element-Wise Lipschitz Activation
by: Daniely, Amit, et al.
Published: (2022)
by: Daniely, Amit, et al.
Published: (2022)
Emulating the Forced Response of Climate Models with Flow Matching
by: Clyne, Graham, et al.
Published: (2026)
by: Clyne, Graham, et al.
Published: (2026)
ArchesWeather & ArchesWeatherGen: a deterministic and generative model for efficient ML weather forecasting
by: Couairon, Guillaume, et al.
Published: (2024)
by: Couairon, Guillaume, et al.
Published: (2024)
Collaborating in Multi-Armed Bandits with Strategic Agents
by: Barnea, Idan, et al.
Published: (2026)
by: Barnea, Idan, et al.
Published: (2026)
Learning Rate Optimization for Deep Neural Networks Using Lipschitz Bandits
by: Priyanka, Padma, et al.
Published: (2024)
by: Priyanka, Padma, et al.
Published: (2024)
Distributed Algorithms for Multi-Agent Multi-Armed Bandits with Collision
by: Zhou, Daoyuan, et al.
Published: (2025)
by: Zhou, Daoyuan, et al.
Published: (2025)
Online Nonsubmodular Optimization with Delayed Feedback in the Bandit Setting
by: Yang, Sifan, et al.
Published: (2025)
by: Yang, Sifan, et al.
Published: (2025)
STIPP: Space-time in situ postprocessing over the French Alps using proper scoring rules
by: Landry, David, et al.
Published: (2026)
by: Landry, David, et al.
Published: (2026)
Multi-Agent Stochastic Bandits Robust to Adversarial Corruptions
by: Ghaffari, Fatemeh, et al.
Published: (2024)
by: Ghaffari, Fatemeh, et al.
Published: (2024)
Adaptive Sample Sharing for Multi Agent Linear Bandits
by: Cherkaoui, Hamza, et al.
Published: (2023)
by: Cherkaoui, Hamza, et al.
Published: (2023)
Lipschitz Dueling Bandits over Continuous Action Spaces
by: Sharma, Mudit, et al.
Published: (2026)
by: Sharma, Mudit, et al.
Published: (2026)
Generative Unsupervised Downscaling of Climate Models via Domain Alignment: Application to Wind Fields
by: Keisler, Julie, et al.
Published: (2026)
by: Keisler, Julie, et al.
Published: (2026)
Improved Regret for Bandit Convex Optimization with Delayed Feedback
by: Wan, Yuanyu, et al.
Published: (2024)
by: Wan, Yuanyu, et al.
Published: (2024)
SerpentFlow: Generative Unpaired Domain Alignment via Shared-Structure Decomposition
by: Keisler, Julie, et al.
Published: (2026)
by: Keisler, Julie, et al.
Published: (2026)
Multi-Objective Multi-Agent Bandits: From Learning Efficiency to Fairness Optimization
by: Wang, John, et al.
Published: (2026)
by: Wang, John, et al.
Published: (2026)
Fair Algorithms with Probing for Multi-Agent Multi-Armed Bandits
by: Xu, Tianyi, et al.
Published: (2025)
by: Xu, Tianyi, et al.
Published: (2025)
PAL: Pluralistic Alignment Framework for Learning from Heterogeneous Preferences
by: Chen, Daiwei, et al.
Published: (2024)
by: Chen, Daiwei, et al.
Published: (2024)
Distributed Multi-Agent Bandits Over Erdős-Rényi Random Networks
by: Liu, Jingyuan, et al.
Published: (2025)
by: Liu, Jingyuan, et al.
Published: (2025)
Multi-Agent Stage-wise Conservative Linear Bandits
by: Afsharrad, Amirhossein, et al.
Published: (2025)
by: Afsharrad, Amirhossein, et al.
Published: (2025)
Cooperative Multi-Agent Constrained Stochastic Linear Bandits
by: Afsharrad, Amirhossein, et al.
Published: (2024)
by: Afsharrad, Amirhossein, et al.
Published: (2024)
An Operator Learning Framework for Spatiotemporal Super-resolution of Scientific Simulations
by: Duruisseaux, Valentin, et al.
Published: (2023)
by: Duruisseaux, Valentin, et al.
Published: (2023)
Multi-Agent Combinatorial-Multi-Armed-Bandit framework for the Submodular Welfare Problem under Bandit Feedback
by: Pokhriyal, Subham, et al.
Published: (2026)
by: Pokhriyal, Subham, et al.
Published: (2026)
Multi-Play Combinatorial Semi-Bandit Problem
by: Nakamura, Shintaro, et al.
Published: (2025)
by: Nakamura, Shintaro, et al.
Published: (2025)
A Survey of Safe Reinforcement Learning and Constrained MDPs: A Technical Survey on Single-Agent and Multi-Agent Safety
by: Kushwaha, Ankita, et al.
Published: (2025)
by: Kushwaha, Ankita, et al.
Published: (2025)
Rising Multi-Armed Bandits with Known Horizons
by: Song, Seockbean, et al.
Published: (2026)
by: Song, Seockbean, et al.
Published: (2026)
Federated Combinatorial Multi-Agent Multi-Armed Bandits
by: Fourati, Fares, et al.
Published: (2024)
by: Fourati, Fares, et al.
Published: (2024)
Similar Items
-
Incentivized Lipschitz Bandits
by: Chakraborty, Sourav, et al.
Published: (2025) -
Flickering Multi-Armed Bandits
by: Chakraborty, Sourav, et al.
Published: (2026) -
A Unified Framework for Locality in Scalable MARL
by: Chakraborty, Sourav, et al.
Published: (2026) -
The Role of Generator Access in Autoregressive Post-Training
by: Rege, Amit Kiran
Published: (2026) -
Data Attribution in Adaptive Learning
by: Rege, Amit Kiran
Published: (2026)