Multi-agent Multi-armed Bandits with Minimum Reward Guarantee Fairness
Fuente:
arXiv
Saved in:
| Main Authors: | Manupriya, Piyushi, Himanshu, Jagarlapudi, SakethaNath, Ghalme, Ganesh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Submodular Framework for Structured-Sparse Optimal Transport
by: Manupriya, Piyushi, et al.
Published: (2024)
by: Manupriya, Piyushi, et al.
Published: (2024)
MMD-Regularized Unbalanced Optimal Transport
by: Manupriya, Piyushi, et al.
Published: (2020)
by: Manupriya, Piyushi, et al.
Published: (2020)
Consistent Optimal Transport with Empirical Conditional Measures
by: Manupriya, Piyushi, et al.
Published: (2023)
by: Manupriya, Piyushi, et al.
Published: (2023)
Learning Counterfactually Fair Models via Improved Generation with Neural Causal Models
by: Kher, Krishn Vishwas, et al.
Published: (2025)
by: Kher, Krishn Vishwas, et al.
Published: (2025)
Finite-Time Guarantees for Multi-Agent Combinatorial Bandits with Nonstationary Rewards
by: Adams, Katherine B., et al.
Published: (2025)
by: Adams, Katherine B., et al.
Published: (2025)
Exponential Convergence of (Stochastic) Gradient Descent for Separable Logistic Regression
by: Kale, Sacchit, et al.
Published: (2026)
by: Kale, Sacchit, et al.
Published: (2026)
Simultaneously Achieving Group Exposure Fairness and Within-Group Meritocracy in Stochastic Bandits
by: Pokhriyal, Subham, et al.
Published: (2024)
by: Pokhriyal, Subham, et al.
Published: (2024)
Lipschitz Dueling Bandits over Continuous Action Spaces
by: Sharma, Mudit, et al.
Published: (2026)
by: Sharma, Mudit, et al.
Published: (2026)
Decentralized Blockchain-based Robust Multi-agent Multi-armed Bandit
by: Xu, Mengfan, et al.
Published: (2024)
by: Xu, Mengfan, et al.
Published: (2024)
Byzantine-Resilient Decentralized Multi-Armed Bandits
by: Zhu, Jingxuan, et al.
Published: (2023)
by: Zhu, Jingxuan, et al.
Published: (2023)
EcoFair-CH-MARL: Scalable Constrained Hierarchical Multi-Agent RL with Real-Time Emission Budgets and Fairness Guarantees
by: Alqithami, Saad
Published: (2026)
by: Alqithami, Saad
Published: (2026)
Tangential Randomization in Linear Bandits (TRAiL): Guaranteed Inference and Regret Bounds
by: Güçlü, Arda, et al.
Published: (2024)
by: Güçlü, Arda, et al.
Published: (2024)
Multi-Agent Stage-wise Conservative Linear Bandits
by: Afsharrad, Amirhossein, et al.
Published: (2025)
by: Afsharrad, Amirhossein, et al.
Published: (2025)
Guaranteeing Control Requirements via Reward Shaping in Reinforcement Learning
by: De Lellis, Francesco, et al.
Published: (2023)
by: De Lellis, Francesco, et al.
Published: (2023)
Decentralized Upper Confidence Bound Algorithms for Homogeneous Multi-Agent Multi-Armed Bandits
by: Zhu, Jingxuan, et al.
Published: (2021)
by: Zhu, Jingxuan, et al.
Published: (2021)
An Analysis of Safety Guarantees in Multi-Task Bayesian Optimization
by: Luebsen, Jannis O., et al.
Published: (2025)
by: Luebsen, Jannis O., et al.
Published: (2025)
Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy
by: Juneja, Ishank, et al.
Published: (2026)
by: Juneja, Ishank, et al.
Published: (2026)
Restless Bandit Problem with Rewards Generated by a Linear Gaussian Dynamical System
by: Gornet, Jonathan, et al.
Published: (2024)
by: Gornet, Jonathan, et al.
Published: (2024)
Meritocratic Fairness in Budgeted Combinatorial Multi-armed Bandits via Shapley Values
by: Sharma, Shradha, et al.
Published: (2026)
by: Sharma, Shradha, et al.
Published: (2026)
Lagrangian Relaxation for Multi-Action Partially Observable Restless Bandits: Heuristic Policies and Indexability
by: Meshram, Rahul, et al.
Published: (2025)
by: Meshram, Rahul, et al.
Published: (2025)
Multi-User mmWave Beam and Rate Adaptation via Combinatorial Satisficing Bandits
by: Özyıldırım, Emre, et al.
Published: (2026)
by: Özyıldırım, Emre, et al.
Published: (2026)
An LP-based Sampling Policy for Multi-Armed Bandits with Side-Observations and Stochastic Availability
by: Soni, Ashutosh, et al.
Published: (2026)
by: Soni, Ashutosh, et al.
Published: (2026)
Probabilistic Safety Guarantee for Stochastic Control Systems Using Average Reward MDPs
by: Omidi, Saber, et al.
Published: (2025)
by: Omidi, Saber, et al.
Published: (2025)
Unsupervised Structural-Counterfactual Generation under Domain Shift
by: Kher, Krishn Vishwas, et al.
Published: (2025)
by: Kher, Krishn Vishwas, et al.
Published: (2025)
Multi-Armed Bandits with Minimum Aggregated Revenue Constraints
by: Yahmed, Ahmed Ben, et al.
Published: (2025)
by: Yahmed, Ahmed Ben, et al.
Published: (2025)
Federated Multi-Agent Deep Reinforcement Learning Approach via Physics-Informed Reward for Multi-Microgrid Energy Management
by: Li, Yuanzheng, et al.
Published: (2022)
by: Li, Yuanzheng, et al.
Published: (2022)
Scalable Data-Driven Reachability Analysis and Control via Koopman Operators with Conformal Coverage Guarantees
by: Nath, Devesh, et al.
Published: (2026)
by: Nath, Devesh, et al.
Published: (2026)
Performance-bounded Online Ensemble Learning Method Based on Multi-armed bandits and Its Applications in Real-time Safety Assessment
by: Hu, Songqiao, et al.
Published: (2025)
by: Hu, Songqiao, et al.
Published: (2025)
Multi-Agent Reinforcement Learning for Offloading Cellular Communications with Cooperating UAVs
by: Mondal, Abhishek, et al.
Published: (2024)
by: Mondal, Abhishek, et al.
Published: (2024)
An Adaptive Method for Contextual Stochastic Multi-armed Bandits with Rewards Generated by a Linear Dynamical System
by: Gornet, Jonathan, et al.
Published: (2024)
by: Gornet, Jonathan, et al.
Published: (2024)
Global Rewards in Multi-Agent Deep Reinforcement Learning for Autonomous Mobility on Demand Systems
by: Hoppe, Heiko, et al.
Published: (2023)
by: Hoppe, Heiko, et al.
Published: (2023)
Multi-CALF: A Policy Combination Approach with Statistical Guarantees
by: Malaniya, Georgiy, et al.
Published: (2025)
by: Malaniya, Georgiy, et al.
Published: (2025)
Adversarial Network Optimization under Bandit Feedback: Maximizing Utility in Non-Stationary Multi-Hop Networks
by: Dai, Yan, et al.
Published: (2024)
by: Dai, Yan, et al.
Published: (2024)
BanditQ: Fair Bandits with Guaranteed Rewards
by: Sinha, Abhishek
Published: (2023)
by: Sinha, Abhishek
Published: (2023)
Scalable Fairness Shaping with LLM-Guided Multi-Agent Reinforcement Learning for Peer-to-Peer Electricity Markets
by: Jadhav, Shrenik, et al.
Published: (2025)
by: Jadhav, Shrenik, et al.
Published: (2025)
Model Predictive Control is almost Optimal for Heterogeneous Restless Multi-armed Bandits
by: Narasimha, Dheeraj, et al.
Published: (2025)
by: Narasimha, Dheeraj, et al.
Published: (2025)
Learning to Sparsify Stochastic Linear Bandits
by: Wang, Zhengmiao, et al.
Published: (2026)
by: Wang, Zhengmiao, et al.
Published: (2026)
Bandit Algorithms for Deep Brain Stimulation
by: Gupta, Arkaprava, et al.
Published: (2026)
by: Gupta, Arkaprava, et al.
Published: (2026)
Faster Q-Learning Algorithms for Restless Bandits
by: Kakarapalli, Parvish, et al.
Published: (2024)
by: Kakarapalli, Parvish, et al.
Published: (2024)
Explore-then-Commit for Nonstationary Linear Bandits with Latent Dynamics
by: Choi, Sunmook, et al.
Published: (2025)
by: Choi, Sunmook, et al.
Published: (2025)
Similar Items
-
Submodular Framework for Structured-Sparse Optimal Transport
by: Manupriya, Piyushi, et al.
Published: (2024) -
MMD-Regularized Unbalanced Optimal Transport
by: Manupriya, Piyushi, et al.
Published: (2020) -
Consistent Optimal Transport with Empirical Conditional Measures
by: Manupriya, Piyushi, et al.
Published: (2023) -
Learning Counterfactually Fair Models via Improved Generation with Neural Causal Models
by: Kher, Krishn Vishwas, et al.
Published: (2025) -
Finite-Time Guarantees for Multi-Agent Combinatorial Bandits with Nonstationary Rewards
by: Adams, Katherine B., et al.
Published: (2025)