The Bandit Whisperer: Communication Learning for Restless Bandits
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Yunfan, Wang, Tonghan, Nagaraj, Dheeraj, Taneja, Aparna, Tambe, Milind |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
by: Behari, Nikhil, et al.
Published: (2024)
by: Behari, Nikhil, et al.
Published: (2024)
Multilinguality in LLM-Designed Reward Functions for Restless Bandits: Effects on Task Performance and Fairness
by: Parthasarathy, Ambreesh, et al.
Published: (2025)
by: Parthasarathy, Ambreesh, et al.
Published: (2025)
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
by: Verma, Shresth, et al.
Published: (2024)
by: Verma, Shresth, et al.
Published: (2024)
Finite-Horizon Single-Pull Restless Bandits: An Efficient Index Policy For Scarce Resource Allocation
by: Xiong, Guojun, et al.
Published: (2025)
by: Xiong, Guojun, et al.
Published: (2025)
Fairness for Workers Who Pull the Arms: An Index Based Policy for Allocation of Restless Bandit Tasks
by: Biswas, Arpita, et al.
Published: (2023)
by: Biswas, Arpita, et al.
Published: (2023)
Towards a Pretrained Model for Restless Bandits via Multi-arm Generalization
by: Zhao, Yunfan, et al.
Published: (2023)
by: Zhao, Yunfan, et al.
Published: (2023)
Combining Diverse Information for Coordinated Action: Stochastic Bandit Algorithms for Heterogeneous Agents
by: Gordon, Lucia, et al.
Published: (2024)
by: Gordon, Lucia, et al.
Published: (2024)
Improving Health Information Access in the World's Largest Maternal Mobile Health Program via Bandit Algorithms
by: Lalan, Arshika, et al.
Published: (2024)
by: Lalan, Arshika, et al.
Published: (2024)
Improving the Prediction of Individual Engagement in Recommendations Using Cognitive Models
by: Seow, Roderick, et al.
Published: (2024)
by: Seow, Roderick, et al.
Published: (2024)
Context in Public Health for Underserved Communities: A Bayesian Approach to Online Restless Bandits
by: Liang, Biyonka, et al.
Published: (2024)
by: Liang, Biyonka, et al.
Published: (2024)
Adversarial Bandit over Bandits: Hierarchical Bandits for Online Configuration Management
by: Avin, Chen, et al.
Published: (2025)
by: Avin, Chen, et al.
Published: (2025)
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
by: Jain, Gauri, et al.
Published: (2024)
by: Jain, Gauri, et al.
Published: (2024)
Decisions and Deployment: The Five-Year SAHELI Project (2020-2025) on Restless Multi-Armed Bandits for Improving Maternal and Child Health
by: Verma, Shresth, et al.
Published: (2026)
by: Verma, Shresth, et al.
Published: (2026)
Anomaly Detection in Networked Bandits
by: Cheng, Xiaotong, et al.
Published: (2025)
by: Cheng, Xiaotong, et al.
Published: (2025)
Cooperative Multi-Agent Constrained Stochastic Linear Bandits
by: Afsharrad, Amirhossein, et al.
Published: (2024)
by: Afsharrad, Amirhossein, et al.
Published: (2024)
Near Optimal Best Arm Identification for Clustered Bandits
by: Yash, et al.
Published: (2025)
by: Yash, et al.
Published: (2025)
Distributed Multi-Task Learning for Stochastic Bandits with Context Distribution and Stage-wise Constraints
by: Lin, Jiabin, et al.
Published: (2024)
by: Lin, Jiabin, et al.
Published: (2024)
Harnessing the Power of Federated Learning in Federated Contextual Bandits
by: Shi, Chengshuai, et al.
Published: (2023)
by: Shi, Chengshuai, et al.
Published: (2023)
Cooperative Multi-Agent Graph Bandits: UCB Algorithm and Regret Analysis
by: Paschalidis, Phevos, et al.
Published: (2024)
by: Paschalidis, Phevos, et al.
Published: (2024)
Decentralized Blockchain-based Robust Multi-agent Multi-armed Bandit
by: Xu, Mengfan, et al.
Published: (2024)
by: Xu, Mengfan, et al.
Published: (2024)
Adaptive Requesting in Decentralized Edge Networks via Non-Stationary Bandits
by: Zhuang, Yi, et al.
Published: (2026)
by: Zhuang, Yi, et al.
Published: (2026)
Lipschitz Dueling Bandits over Continuous Action Spaces
by: Sharma, Mudit, et al.
Published: (2026)
by: Sharma, Mudit, et al.
Published: (2026)
Generative AI Against Poaching: Latent Composite Flow Matching for Wildlife Conservation
by: Kong, Lingkai, et al.
Published: (2025)
by: Kong, Lingkai, et al.
Published: (2025)
The Nah Bandit: Modeling User Non-compliance in Recommendation Systems
by: Zhou, Tianyue, et al.
Published: (2024)
by: Zhou, Tianyue, et al.
Published: (2024)
In-Context Curiosity: Distilling Exploration for Decision-Pretrained Transformers on Bandit Tasks
by: Yang, Huitao, et al.
Published: (2025)
by: Yang, Huitao, et al.
Published: (2025)
Meritocratic Fairness in Budgeted Combinatorial Multi-armed Bandits via Shapley Values
by: Sharma, Shradha, et al.
Published: (2026)
by: Sharma, Shradha, et al.
Published: (2026)
Dynamic Matching Bandit For Two-Sided Online Markets
by: Li, Yuantong, et al.
Published: (2022)
by: Li, Yuantong, et al.
Published: (2022)
Multi-Agent Bandit Learning through Heterogeneous Action Erasure Channels
by: Hanna, Osama A., et al.
Published: (2023)
by: Hanna, Osama A., et al.
Published: (2023)
Bayesian Collaborative Bandits with Thompson Sampling for Improved Outreach in Maternal Health Program
by: Dasgupta, Arpan, et al.
Published: (2024)
by: Dasgupta, Arpan, et al.
Published: (2024)
Federated Combinatorial Multi-Agent Multi-Armed Bandits
by: Fourati, Fares, et al.
Published: (2024)
by: Fourati, Fares, et al.
Published: (2024)
Heterogeneous Multi-Agent Bandits with Parsimonious Hints
by: Mirfakhar, Amirmahdi, et al.
Published: (2025)
by: Mirfakhar, Amirmahdi, et al.
Published: (2025)
Procedural Fairness in Multi-Agent Bandits
by: Caiata, Joshua, et al.
Published: (2026)
by: Caiata, Joshua, et al.
Published: (2026)
Simultaneously Achieving Group Exposure Fairness and Within-Group Meritocracy in Stochastic Bandits
by: Pokhriyal, Subham, et al.
Published: (2024)
by: Pokhriyal, Subham, et al.
Published: (2024)
Faster Q-Learning Algorithms for Restless Bandits
by: Kakarapalli, Parvish, et al.
Published: (2024)
by: Kakarapalli, Parvish, et al.
Published: (2024)
Principal-Agent Reinforcement Learning: Orchestrating AI Agents with Contracts
by: Ivanov, Dima, et al.
Published: (2024)
by: Ivanov, Dima, et al.
Published: (2024)
Communications-Incentivized Collaborative Reasoning in NetGPT through Agentic Reinforcement Learning
by: Yu, Xiaoxue, et al.
Published: (2026)
by: Yu, Xiaoxue, et al.
Published: (2026)
A Survey of Multi-Agent Deep Reinforcement Learning with Communication
by: Zhu, Changxi, et al.
Published: (2022)
by: Zhu, Changxi, et al.
Published: (2022)
Feature-based Federated Transfer Learning: Communication Efficiency, Robustness and Privacy
by: Wang, Feng, et al.
Published: (2024)
by: Wang, Feng, et al.
Published: (2024)
Implicit Repair with Reinforcement Learning in Emergent Communication
by: Vital, Fábio, et al.
Published: (2025)
by: Vital, Fábio, et al.
Published: (2025)
Fully Independent Communication in Multi-Agent Reinforcement Learning
by: Pina, Rafael, et al.
Published: (2024)
by: Pina, Rafael, et al.
Published: (2024)
Similar Items
-
A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
by: Behari, Nikhil, et al.
Published: (2024) -
Multilinguality in LLM-Designed Reward Functions for Restless Bandits: Effects on Task Performance and Fairness
by: Parthasarathy, Ambreesh, et al.
Published: (2025) -
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
by: Verma, Shresth, et al.
Published: (2024) -
Finite-Horizon Single-Pull Restless Bandits: An Efficient Index Policy For Scarce Resource Allocation
by: Xiong, Guojun, et al.
Published: (2025) -
Fairness for Workers Who Pull the Arms: An Index Based Policy for Allocation of Restless Bandit Tasks
by: Biswas, Arpita, et al.
Published: (2023)