Navigating the Social Welfare Frontier: Portfolios for Multi-objective Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Cheol Woo, Moondra, Jai, Verma, Shresth, Pollack, Madeleine, Kong, Lingkai, Tambe, Milind, Gupta, Swati |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
by: Verma, Shresth, et al.
Published: (2024)
by: Verma, Shresth, et al.
Published: (2024)
Lightweight Robust Direct Preference Optimization
by: Kim, Cheol Woo, et al.
Published: (2025)
by: Kim, Cheol Woo, et al.
Published: (2025)
Preference Robustness for DPO with Applications to Public Health
by: Kim, Cheol Woo, et al.
Published: (2025)
by: Kim, Cheol Woo, et al.
Published: (2025)
Many Preferences, Few Policies: Towards Scalable Language Model Personalization
by: Kim, Cheol Woo, et al.
Published: (2026)
by: Kim, Cheol Woo, et al.
Published: (2026)
Improved Regret Guarantees for Online Mirror Descent using a Portfolio of Mirror Maps
by: Gupta, Swati, et al.
Published: (2026)
by: Gupta, Swati, et al.
Published: (2026)
Generative AI for Social Impact
by: Kong, Lingkai, et al.
Published: (2026)
by: Kong, Lingkai, et al.
Published: (2026)
Why Global LLM Leaderboards Are Misleading: Small Portfolios for Heterogeneous Supervised ML
by: Moondra, Jai, et al.
Published: (2026)
by: Moondra, Jai, et al.
Published: (2026)
Provably Small Portfolios for Multiobjective Optimization with Application to Subsidized Facility Location
by: Gupta, Swati, et al.
Published: (2025)
by: Gupta, Swati, et al.
Published: (2025)
Balancing Notions of Equity: Trade-offs Between Fair Portfolio Sizes and Achievable Guarantees
by: Gupta, Swati, et al.
Published: (2023)
by: Gupta, Swati, et al.
Published: (2023)
Composite Flow Matching for Reinforcement Learning with Shifted-Dynamics Data
by: Kong, Lingkai, et al.
Published: (2025)
by: Kong, Lingkai, et al.
Published: (2025)
Decisions and Deployment: The Five-Year SAHELI Project (2020-2025) on Restless Multi-Armed Bandits for Improving Maternal and Child Health
by: Verma, Shresth, et al.
Published: (2026)
by: Verma, Shresth, et al.
Published: (2026)
LLM-based Agent Simulation for Maternal Health Interventions: Uncertainty Estimation and Decision-focused Evaluation
by: Martinson, Sarah, et al.
Published: (2025)
by: Martinson, Sarah, et al.
Published: (2025)
Which $L_p$ norm is the fairest? Approximations for fair facility location across all "$p$"
by: Gupta, Swati, et al.
Published: (2022)
by: Gupta, Swati, et al.
Published: (2022)
Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions
by: Kong, Lingkai, et al.
Published: (2026)
by: Kong, Lingkai, et al.
Published: (2026)
Reward Shaping for Inference-Time Alignment: A Stackelberg Game Perspective
by: Wang, Haichuan, et al.
Published: (2026)
by: Wang, Haichuan, et al.
Published: (2026)
What is the Right Notion of Distance between Predict-then-Optimize Tasks?
by: Rodriguez-Diaz, Paula, et al.
Published: (2024)
by: Rodriguez-Diaz, Paula, et al.
Published: (2024)
Promise of Graph Sparsification and Decomposition for Noise Reduction in QAOA: Analysis for Trapped-Ion Compilations
by: Moondra, Jai, et al.
Published: (2024)
by: Moondra, Jai, et al.
Published: (2024)
Generative AI Against Poaching: Latent Composite Flow Matching for Wildlife Conservation
by: Kong, Lingkai, et al.
Published: (2025)
by: Kong, Lingkai, et al.
Published: (2025)
Robust Optimization with Diffusion Models for Green Security
by: Kong, Lingkai, et al.
Published: (2025)
by: Kong, Lingkai, et al.
Published: (2025)
Incentive-Aware AI Safety via Strategic Resource Allocation: A Stackelberg Security Games Perspective
by: Kim, Cheol Woo, et al.
Published: (2026)
by: Kim, Cheol Woo, et al.
Published: (2026)
Rule-Bottleneck Reinforcement Learning: Joint Explanation and Decision Optimization for Resource Allocation with Language Agents
by: Tec, Mauricio, et al.
Published: (2025)
by: Tec, Mauricio, et al.
Published: (2025)
Improving Health Information Access in the World's Largest Maternal Mobile Health Program via Bandit Algorithms
by: Lalan, Arshika, et al.
Published: (2024)
by: Lalan, Arshika, et al.
Published: (2024)
Reinforcement learning with combinatorial actions for coupled restless bandits
by: Xu, Lily, et al.
Published: (2025)
by: Xu, Lily, et al.
Published: (2025)
Analyzing Cost-Sensitive Surrogate Losses via $\mathcal{H}$-calibration
by: Shah, Sanket, et al.
Published: (2025)
by: Shah, Sanket, et al.
Published: (2025)
Efficient Public Health Intervention Planning Using Decomposition-Based Decision-Focused Learning
by: Shah, Sanket, et al.
Published: (2024)
by: Shah, Sanket, et al.
Published: (2024)
Combining Diverse Information for Coordinated Action: Stochastic Bandit Algorithms for Heterogeneous Agents
by: Gordon, Lucia, et al.
Published: (2024)
by: Gordon, Lucia, et al.
Published: (2024)
Equitably allocating wildfire resilience investments for power grids: The curse of aggregation and vulnerability indices
by: Pollack, Madeleine, et al.
Published: (2024)
by: Pollack, Madeleine, et al.
Published: (2024)
Adaptive Frontier Exploration on Graphs with Applications to Network-Based Disease Testing
by: Choo, Davin, et al.
Published: (2025)
by: Choo, Davin, et al.
Published: (2025)
Welfare and Fairness in Multi-objective Reinforcement Learning
by: Fan, Zimeng, et al.
Published: (2022)
by: Fan, Zimeng, et al.
Published: (2022)
Optimal Control of Multiclass Fluid Queueing Networks: A Machine Learning Approach
by: Bertsimas, Dimitris, et al.
Published: (2023)
by: Bertsimas, Dimitris, et al.
Published: (2023)
Advancing Investment Frontiers: Industry-grade Deep Reinforcement Learning for Portfolio Optimization
by: Ndikum, Philip, et al.
Published: (2024)
by: Ndikum, Philip, et al.
Published: (2024)
Policy-Embedded Graph Expansion: Networked HIV Testing with Diffusion-Driven Network Samples
by: Kangaslahti, Akseli, et al.
Published: (2026)
by: Kangaslahti, Akseli, et al.
Published: (2026)
Optimal Control of Fluid Restless Multi-armed Bandits: A Machine Learning Approach
by: Bertsimas, Dimitris, et al.
Published: (2025)
by: Bertsimas, Dimitris, et al.
Published: (2025)
The Bandit Whisperer: Communication Learning for Restless Bandits
by: Zhao, Yunfan, et al.
Published: (2024)
by: Zhao, Yunfan, et al.
Published: (2024)
A Machine Learning Approach to Two-Stage Adaptive Robust Optimization
by: Bertsimas, Dimitris, et al.
Published: (2023)
by: Bertsimas, Dimitris, et al.
Published: (2023)
On Diffusion Models for Multi-Agent Partial Observability: Shared Attractors, Error Bounds, and Composite Flow
by: Wang, Tonghan, et al.
Published: (2024)
by: Wang, Tonghan, et al.
Published: (2024)
Quantitative Convergences of Lie Group Momentum Optimizers
by: Kong, Lingkai, et al.
Published: (2024)
by: Kong, Lingkai, et al.
Published: (2024)
Convergence of Kinetic Langevin Monte Carlo on Lie groups
by: Kong, Lingkai, et al.
Published: (2024)
by: Kong, Lingkai, et al.
Published: (2024)
Leaving the Nest: Going Beyond Local Loss Functions for Predict-Then-Optimize
by: Shah, Sanket, et al.
Published: (2023)
by: Shah, Sanket, et al.
Published: (2023)
Dual-Mandate Patrols: Multi-Armed Bandits for Green Security
by: Xu, Lily, et al.
Published: (2020)
by: Xu, Lily, et al.
Published: (2020)
Similar Items
-
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
by: Verma, Shresth, et al.
Published: (2024) -
Lightweight Robust Direct Preference Optimization
by: Kim, Cheol Woo, et al.
Published: (2025) -
Preference Robustness for DPO with Applications to Public Health
by: Kim, Cheol Woo, et al.
Published: (2025) -
Many Preferences, Few Policies: Towards Scalable Language Model Personalization
by: Kim, Cheol Woo, et al.
Published: (2026) -
Improved Regret Guarantees for Online Mirror Descent using a Portfolio of Mirror Maps
by: Gupta, Swati, et al.
Published: (2026)