Multilinguality in LLM-Designed Reward Functions for Restless Bandits: Effects on Task Performance and Fairness
Fuente:
arXiv
Saved in:
| Main Authors: | Parthasarathy, Ambreesh, Subramanian, Chandrasekar, Senrayan, Ganesh, Adappanavar, Shreyash, Taneja, Aparna, Ravindran, Balaraman, Tambe, Milind |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Bandit Whisperer: Communication Learning for Restless Bandits
by: Zhao, Yunfan, et al.
Published: (2024)
by: Zhao, Yunfan, et al.
Published: (2024)
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
by: Verma, Shresth, et al.
Published: (2024)
by: Verma, Shresth, et al.
Published: (2024)
A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
by: Behari, Nikhil, et al.
Published: (2024)
by: Behari, Nikhil, et al.
Published: (2024)
Fairness for Workers Who Pull the Arms: An Index Based Policy for Allocation of Restless Bandit Tasks
by: Biswas, Arpita, et al.
Published: (2023)
by: Biswas, Arpita, et al.
Published: (2023)
Finite-Horizon Single-Pull Restless Bandits: An Efficient Index Policy For Scarce Resource Allocation
by: Xiong, Guojun, et al.
Published: (2025)
by: Xiong, Guojun, et al.
Published: (2025)
Improving Health Information Access in the World's Largest Maternal Mobile Health Program via Bandit Algorithms
by: Lalan, Arshika, et al.
Published: (2024)
by: Lalan, Arshika, et al.
Published: (2024)
mFARM: Towards Multi-Faceted Fairness Assessment based on HARMs in Clinical Decision Support
by: Adappanavar, Shreyash, et al.
Published: (2025)
by: Adappanavar, Shreyash, et al.
Published: (2025)
Combining Diverse Information for Coordinated Action: Stochastic Bandit Algorithms for Heterogeneous Agents
by: Gordon, Lucia, et al.
Published: (2024)
by: Gordon, Lucia, et al.
Published: (2024)
Context in Public Health for Underserved Communities: A Bayesian Approach to Online Restless Bandits
by: Liang, Biyonka, et al.
Published: (2024)
by: Liang, Biyonka, et al.
Published: (2024)
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
by: Jain, Gauri, et al.
Published: (2024)
by: Jain, Gauri, et al.
Published: (2024)
Decisions and Deployment: The Five-Year SAHELI Project (2020-2025) on Restless Multi-Armed Bandits for Improving Maternal and Child Health
by: Verma, Shresth, et al.
Published: (2026)
by: Verma, Shresth, et al.
Published: (2026)
MABL: Bi-Level Latent-Variable World Model for Sample-Efficient Multi-Agent Reinforcement Learning
by: Venugopal, Aravind, et al.
Published: (2023)
by: Venugopal, Aravind, et al.
Published: (2023)
Towards a Pretrained Model for Restless Bandits via Multi-arm Generalization
by: Zhao, Yunfan, et al.
Published: (2023)
by: Zhao, Yunfan, et al.
Published: (2023)
Improving the Prediction of Individual Engagement in Recommendations Using Cognitive Models
by: Seow, Roderick, et al.
Published: (2024)
by: Seow, Roderick, et al.
Published: (2024)
Participatory Approaches in AI Development and Governance: Case Studies
by: Parthasarathy, Ambreesh, et al.
Published: (2024)
by: Parthasarathy, Ambreesh, et al.
Published: (2024)
Simultaneously Achieving Group Exposure Fairness and Within-Group Meritocracy in Stochastic Bandits
by: Pokhriyal, Subham, et al.
Published: (2024)
by: Pokhriyal, Subham, et al.
Published: (2024)
Learning to Coordinate Under Threshold Rewards: A Cooperative Multi-Agent Bandit Framework
by: Ledford, Michael, et al.
Published: (2025)
by: Ledford, Michael, et al.
Published: (2025)
ReSo: A Reward-driven Self-organizing LLM-based Multi-Agent System for Reasoning Tasks
by: Zhou, Heng, et al.
Published: (2025)
by: Zhou, Heng, et al.
Published: (2025)
Participatory Approaches in AI Development and Governance: A Principled Approach
by: Parthasarathy, Ambreesh, et al.
Published: (2024)
by: Parthasarathy, Ambreesh, et al.
Published: (2024)
Generative AI Against Poaching: Latent Composite Flow Matching for Wildlife Conservation
by: Kong, Lingkai, et al.
Published: (2025)
by: Kong, Lingkai, et al.
Published: (2025)
VORTEX: Aligning Task Utility and Human Preferences through LLM-Guided Reward Shaping
by: Xiong, Guojun, et al.
Published: (2025)
by: Xiong, Guojun, et al.
Published: (2025)
Lipschitz Dueling Bandits over Continuous Action Spaces
by: Sharma, Mudit, et al.
Published: (2026)
by: Sharma, Mudit, et al.
Published: (2026)
Bayesian Collaborative Bandits with Thompson Sampling for Improved Outreach in Maternal Health Program
by: Dasgupta, Arpan, et al.
Published: (2024)
by: Dasgupta, Arpan, et al.
Published: (2024)
Relational Weight Optimization for Enhancing Team Performance in Multi-Agent Multi-Armed Bandits
by: Kotturu, Monish Reddy, et al.
Published: (2024)
by: Kotturu, Monish Reddy, et al.
Published: (2024)
Efficient Public Health Intervention Planning Using Decomposition-Based Decision-Focused Learning
by: Shah, Sanket, et al.
Published: (2024)
by: Shah, Sanket, et al.
Published: (2024)
Towards Foundation-model-based Multiagent System to Accelerate AI for Social Impact
by: Zhao, Yunfan, et al.
Published: (2024)
by: Zhao, Yunfan, et al.
Published: (2024)
Logic-based Task Representation and Reward Shaping in Multiagent Reinforcement Learning
by: Doshi, Nishant
Published: (2025)
by: Doshi, Nishant
Published: (2025)
Procedural Fairness in Multi-Agent Bandits
by: Caiata, Joshua, et al.
Published: (2026)
by: Caiata, Joshua, et al.
Published: (2026)
Adversarial Bandit over Bandits: Hierarchical Bandits for Online Configuration Management
by: Avin, Chen, et al.
Published: (2025)
by: Avin, Chen, et al.
Published: (2025)
Meritocratic Fairness in Budgeted Combinatorial Multi-armed Bandits via Shapley Values
by: Sharma, Shradha, et al.
Published: (2026)
by: Sharma, Shradha, et al.
Published: (2026)
Distributed Multi-Task Learning for Stochastic Bandits with Context Distribution and Stage-wise Constraints
by: Lin, Jiabin, et al.
Published: (2024)
by: Lin, Jiabin, et al.
Published: (2024)
Beyond Task Completion: An Assessment Framework for Evaluating Agentic AI Systems
by: Akshathala, Sreemaee, et al.
Published: (2025)
by: Akshathala, Sreemaee, et al.
Published: (2025)
In-Context Curiosity: Distilling Exploration for Decision-Pretrained Transformers on Bandit Tasks
by: Yang, Huitao, et al.
Published: (2025)
by: Yang, Huitao, et al.
Published: (2025)
LLM-based Agent Simulation for Maternal Health Interventions: Uncertainty Estimation and Decision-focused Evaluation
by: Martinson, Sarah, et al.
Published: (2025)
by: Martinson, Sarah, et al.
Published: (2025)
Interactional Fairness in LLM Multi-Agent Systems: An Evaluation Framework
by: Binkyte, Ruta
Published: (2025)
by: Binkyte, Ruta
Published: (2025)
Learning Reward Machines in Cooperative Multi-Agent Tasks
by: Ardon, Leo, et al.
Published: (2023)
by: Ardon, Leo, et al.
Published: (2023)
Sustainable Multi-Agent Crowdsourcing via Physics-Informed Bandits
by: Banerjee, Chayan
Published: (2026)
by: Banerjee, Chayan
Published: (2026)
Logic-Based Verification of Task Allocation for LLM-Enabled Multi-Agent Manufacturing Systems
by: Lim, Jonghan, et al.
Published: (2026)
by: Lim, Jonghan, et al.
Published: (2026)
Learnings from Implementation of a BDI Agent-based Battery-less Wireless Sensor
by: Ramanathan, Ganesh, et al.
Published: (2024)
by: Ramanathan, Ganesh, et al.
Published: (2024)
Anomaly Detection in Networked Bandits
by: Cheng, Xiaotong, et al.
Published: (2025)
by: Cheng, Xiaotong, et al.
Published: (2025)
Similar Items
-
The Bandit Whisperer: Communication Learning for Restless Bandits
by: Zhao, Yunfan, et al.
Published: (2024) -
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
by: Verma, Shresth, et al.
Published: (2024) -
A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
by: Behari, Nikhil, et al.
Published: (2024) -
Fairness for Workers Who Pull the Arms: An Index Based Policy for Allocation of Restless Bandit Tasks
by: Biswas, Arpita, et al.
Published: (2023) -
Finite-Horizon Single-Pull Restless Bandits: An Efficient Index Policy For Scarce Resource Allocation
by: Xiong, Guojun, et al.
Published: (2025)