A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Behari, Nikhil, Zhang, Edwin, Zhao, Yunfan, Taneja, Aparna, Nagaraj, Dheeraj, Tambe, Milind |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Bandit Whisperer: Communication Learning for Restless Bandits
von: Zhao, Yunfan, et al.
Veröffentlicht: (2024)
von: Zhao, Yunfan, et al.
Veröffentlicht: (2024)
Towards a Pretrained Model for Restless Bandits via Multi-arm Generalization
von: Zhao, Yunfan, et al.
Veröffentlicht: (2023)
von: Zhao, Yunfan, et al.
Veröffentlicht: (2023)
Multilinguality in LLM-Designed Reward Functions for Restless Bandits: Effects on Task Performance and Fairness
von: Parthasarathy, Ambreesh, et al.
Veröffentlicht: (2025)
von: Parthasarathy, Ambreesh, et al.
Veröffentlicht: (2025)
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
von: Verma, Shresth, et al.
Veröffentlicht: (2024)
von: Verma, Shresth, et al.
Veröffentlicht: (2024)
Fairness for Workers Who Pull the Arms: An Index Based Policy for Allocation of Restless Bandit Tasks
von: Biswas, Arpita, et al.
Veröffentlicht: (2023)
von: Biswas, Arpita, et al.
Veröffentlicht: (2023)
Finite-Horizon Single-Pull Restless Bandits: An Efficient Index Policy For Scarce Resource Allocation
von: Xiong, Guojun, et al.
Veröffentlicht: (2025)
von: Xiong, Guojun, et al.
Veröffentlicht: (2025)
Improving Health Information Access in the World's Largest Maternal Mobile Health Program via Bandit Algorithms
von: Lalan, Arshika, et al.
Veröffentlicht: (2024)
von: Lalan, Arshika, et al.
Veröffentlicht: (2024)
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
von: Jain, Gauri, et al.
Veröffentlicht: (2024)
von: Jain, Gauri, et al.
Veröffentlicht: (2024)
Decisions and Deployment: The Five-Year SAHELI Project (2020-2025) on Restless Multi-Armed Bandits for Improving Maternal and Child Health
von: Verma, Shresth, et al.
Veröffentlicht: (2026)
von: Verma, Shresth, et al.
Veröffentlicht: (2026)
Improving the Prediction of Individual Engagement in Recommendations Using Cognitive Models
von: Seow, Roderick, et al.
Veröffentlicht: (2024)
von: Seow, Roderick, et al.
Veröffentlicht: (2024)
DLM: Unified Decision Language Models for Offline Multi-Agent Sequential Decision Making
von: Zhang, Zhuohui, et al.
Veröffentlicht: (2026)
von: Zhang, Zhuohui, et al.
Veröffentlicht: (2026)
Combining Diverse Information for Coordinated Action: Stochastic Bandit Algorithms for Heterogeneous Agents
von: Gordon, Lucia, et al.
Veröffentlicht: (2024)
von: Gordon, Lucia, et al.
Veröffentlicht: (2024)
Context in Public Health for Underserved Communities: A Bayesian Approach to Online Restless Bandits
von: Liang, Biyonka, et al.
Veröffentlicht: (2024)
von: Liang, Biyonka, et al.
Veröffentlicht: (2024)
Relational Weight Optimization for Enhancing Team Performance in Multi-Agent Multi-Armed Bandits
von: Kotturu, Monish Reddy, et al.
Veröffentlicht: (2024)
von: Kotturu, Monish Reddy, et al.
Veröffentlicht: (2024)
Federated Combinatorial Multi-Agent Multi-Armed Bandits
von: Fourati, Fares, et al.
Veröffentlicht: (2024)
von: Fourati, Fares, et al.
Veröffentlicht: (2024)
Lark: Biologically Inspired Neuroevolution for Multi-Stakeholder LLM Agents
von: Tanugula, Rikhil, et al.
Veröffentlicht: (2025)
von: Tanugula, Rikhil, et al.
Veröffentlicht: (2025)
Near Optimal Best Arm Identification for Clustered Bandits
von: Yash, et al.
Veröffentlicht: (2025)
von: Yash, et al.
Veröffentlicht: (2025)
Collaborative Multi-Agent Heterogeneous Multi-Armed Bandits
von: Chawla, Ronshee, et al.
Veröffentlicht: (2023)
von: Chawla, Ronshee, et al.
Veröffentlicht: (2023)
Dynamic Task Adaptation for Multi-Robot Manufacturing Systems with Large Language Models
von: Lim, Jonghan, et al.
Veröffentlicht: (2025)
von: Lim, Jonghan, et al.
Veröffentlicht: (2025)
Grounded Predictions of Teamwork as a One-Shot Game: A Multiagent Multi-Armed Bandits Approach
von: Gómez, Alejandra López de Aberasturi, et al.
Veröffentlicht: (2024)
von: Gómez, Alejandra López de Aberasturi, et al.
Veröffentlicht: (2024)
In-Context Curiosity: Distilling Exploration for Decision-Pretrained Transformers on Bandit Tasks
von: Yang, Huitao, et al.
Veröffentlicht: (2025)
von: Yang, Huitao, et al.
Veröffentlicht: (2025)
From General Relation Patterns to Task-Specific Decision-Making in Continual Multi-Agent Coordination
von: Yao, Chang, et al.
Veröffentlicht: (2025)
von: Yao, Chang, et al.
Veröffentlicht: (2025)
Efficient Public Health Intervention Planning Using Decomposition-Based Decision-Focused Learning
von: Shah, Sanket, et al.
Veröffentlicht: (2024)
von: Shah, Sanket, et al.
Veröffentlicht: (2024)
Distributed Multi-Task Learning for Stochastic Bandits with Context Distribution and Stage-wise Constraints
von: Lin, Jiabin, et al.
Veröffentlicht: (2024)
von: Lin, Jiabin, et al.
Veröffentlicht: (2024)
Sustainable Multi-Agent Crowdsourcing via Physics-Informed Bandits
von: Banerjee, Chayan
Veröffentlicht: (2026)
von: Banerjee, Chayan
Veröffentlicht: (2026)
A Multi-Agent Reinforcement Learning Framework for Public Health Decision Analysis
von: Sharma, Dinesh, et al.
Veröffentlicht: (2023)
von: Sharma, Dinesh, et al.
Veröffentlicht: (2023)
Dynamic Strategy Adaptation in Multi-Agent Environments with Large Language Models
von: Mallampati, Shaurya, et al.
Veröffentlicht: (2025)
von: Mallampati, Shaurya, et al.
Veröffentlicht: (2025)
Generative AI Against Poaching: Latent Composite Flow Matching for Wildlife Conservation
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
COLA: A Scalable Multi-Agent Framework For Windows UI Task Automation
von: Zhao, Di, et al.
Veröffentlicht: (2025)
von: Zhao, Di, et al.
Veröffentlicht: (2025)
UMBRELLA: Uncertainty-aware Multi-robot Reactive Coordination under Dynamic Temporal Logic Tasks
von: Zhao, Qisheng, et al.
Veröffentlicht: (2026)
von: Zhao, Qisheng, et al.
Veröffentlicht: (2026)
Learning to Coordinate Under Threshold Rewards: A Cooperative Multi-Agent Bandit Framework
von: Ledford, Michael, et al.
Veröffentlicht: (2025)
von: Ledford, Michael, et al.
Veröffentlicht: (2025)
Towards Foundation-model-based Multiagent System to Accelerate AI for Social Impact
von: Zhao, Yunfan, et al.
Veröffentlicht: (2024)
von: Zhao, Yunfan, et al.
Veröffentlicht: (2024)
Multi-Agent Synchronization Tasks
von: Fernandez, Rolando, et al.
Veröffentlicht: (2024)
von: Fernandez, Rolando, et al.
Veröffentlicht: (2024)
MF-LLM: Simulating Population Decision Dynamics via a Mean-Field Large Language Model Framework
von: Mi, Qirui, et al.
Veröffentlicht: (2025)
von: Mi, Qirui, et al.
Veröffentlicht: (2025)
Learning Policies for Dynamic Coalition Formation in Multi-Robot Task Allocation
von: Bezerra, Lucas C. D., et al.
Veröffentlicht: (2024)
von: Bezerra, Lucas C. D., et al.
Veröffentlicht: (2024)
AGCo-MATA: Air-Ground Collaborative Multi-Agent Task Allocation in Mobile Crowdsensing
von: Shao, Tianhao, et al.
Veröffentlicht: (2025)
von: Shao, Tianhao, et al.
Veröffentlicht: (2025)
Adversarial Bandit over Bandits: Hierarchical Bandits for Online Configuration Management
von: Avin, Chen, et al.
Veröffentlicht: (2025)
von: Avin, Chen, et al.
Veröffentlicht: (2025)
Debate or Vote: Which Yields Better Decisions in Multi-Agent Large Language Models?
von: Choi, Hyeong Kyu, et al.
Veröffentlicht: (2025)
von: Choi, Hyeong Kyu, et al.
Veröffentlicht: (2025)
Task-Aware LLM Council with Adaptive Decision Pathways for Decision Support
von: Zhu, Wei, et al.
Veröffentlicht: (2026)
von: Zhu, Wei, et al.
Veröffentlicht: (2026)
Decentralized Blockchain-based Robust Multi-agent Multi-armed Bandit
von: Xu, Mengfan, et al.
Veröffentlicht: (2024)
von: Xu, Mengfan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
The Bandit Whisperer: Communication Learning for Restless Bandits
von: Zhao, Yunfan, et al.
Veröffentlicht: (2024) -
Towards a Pretrained Model for Restless Bandits via Multi-arm Generalization
von: Zhao, Yunfan, et al.
Veröffentlicht: (2023) -
Multilinguality in LLM-Designed Reward Functions for Restless Bandits: Effects on Task Performance and Fairness
von: Parthasarathy, Ambreesh, et al.
Veröffentlicht: (2025) -
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
von: Verma, Shresth, et al.
Veröffentlicht: (2024) -
Fairness for Workers Who Pull the Arms: An Index Based Policy for Allocation of Restless Bandit Tasks
von: Biswas, Arpita, et al.
Veröffentlicht: (2023)